Baseten Inference Runtime

Infrastructure service

Named inference runtime (automatic TensorRT/SGLang/vLLM builds, speculative decoding, custom kernel fusion, KV-cache optimisation) that underlies Baseten's Dedicated Inference, Model APIs and Training products.

Find alternatives to Baseten Inference Runtime

TypeInfrastructure service
RoleEngine
AvailabilitySold
DeploymentSaaS
Intended forDeveloper, Enterprise
SecurityNot recorded
StatusActive
Sold withinBaseten — Included with the parentOfficial documentation · 14 Sep 2026
PricingNot recorded
Sources—

Not sold on its own; usage draws from whichever Baseten product (Dedicated Inference, Model APIs, Training) it runs underneath. Whether a recurring free tier exists is not stated either way.

What it does

baseten.co

Sources

Description

Official documentation · 14 Sep 2026

What it does

Official documentation · 14 Sep 2026

Sold within

Official documentation · 14 Sep 2026

Page states it is 'the technical foundation... supporting Baseten's broader platform offerings -- including Dedicated Inference, Model APIs, and Training'.

Name

Official documentation · 14 Sep 2026

URL

Official documentation · 14 Sep 2026

Also from Baseten

Open Baseten in the directory

Something wrong here? Send a correction — quote this product id: baseten-inference-runtime.