Elastic Inference Service

Model API

Hosted inference endpoint that runs Elastic-managed LLMs, the ELSER sparse-embedding model and third-party embedding models for ingest, search and chat without provisioning ML nodes in a customer's own Elasticsearch deployment.

Find alternatives to Elastic Inference Service

TypeModel API
RoleHosted model API
AvailabilitySold
DeploymentSaaS
Intended forDeveloperSelf-reported · 8 Sep 2026
SecurityNot recorded
StatusActive
Sold withinElastic Cloud — Paid add-on to the parentOfficial documentation · 29 Sep 2026
Sources1 of 1 field

Billed per million tokens for every model on the service, on top of the Elastic Cloud subscription it is reached from. No marketing page exists — the docs name it and state the billing unit, admitted under the docs-plus-price ruling of 2026-09-04.

What it does

elastic.co

How it compares

Sources

Pricing

Official documentation · 4 Sep 2026

"All models on EIS incur a charge per million tokens … EIS is billed per million tokens used". A rate card per model is not printed on this page.

Description

Official documentation · 4 Sep 2026

Docs: EIS lets you "use machine learning models for ingest, search, and chat independently of your Elasticsearch infrastructure"; lists Elastic Managed LLMs, ELSER on EIS and jina-embeddings-v3.

Sold within

Official documentation · 29 Sep 2026

The docs call it "a global service on Elastic Cloud", metered separately from ML-node VCUs: "All models on EIS incur a charge per million tokens."

Also from Elastic

Open Elastic in the directory

Something wrong here? Send a correction — quote this product id: elastic-inference-service.