Baseten
Deploys and serves machine-learning models as production endpoints on managed GPU infrastructure.
New accounts get credits; dedicated deployments and training bill per minute, and Pro and Enterprise tiers are quoted.
What it does
- Model inference
- Model training
How it compares
-
Fireworks AI
Fireworks AI's closest competitors are managed inference companies such as Together AI and Baseten.
Research report · 19 Sep 2026 -
Modal
Direct competitors include Baseten, which raised $150M Series D and focuses on mission-critical inference with dedicated deployments
Research report · 19 Sep 2026
Sources
- Pricing
-
New accounts get credits; dedicated deployments and training bill per minute, and Pro and Enterprise tiers are quoted.
- Sold within
-
The Basic plan is listed at "$0 per month, pay as you go", and "new Baseten accounts come with credits so you can get to know the UI and experiment with deployments for free." Pro and Enterprise are quoted ("Get a quote").
Also from Baseten
-
Baseten Model APIs Model API
Pre-optimised hosted endpoints for open-source frontier models, called without deploying or managing a deployment first.
-
Baseten Training Platform
Trains and fine-tunes models with reinforcement learning through the Loops SDK, deploying the result onto Baseten's inference stack.
-
Baseten Frontier Gateway API service
Production-grade API launch platform for model labs: takes a lab's model from research to a reliable, scalable, white-labelled API in days, distinct from Baseten's Distribution Platform (model marketplace listing).
-
Baseten Inference Runtime Infrastructure service
Named inference runtime (automatic TensorRT/SGLang/vLLM builds, speculative decoding, custom kernel fusion, KV-cache optimisation) that underlies Baseten's Dedicated Inference, Model APIs and Training products.
Something wrong here? Send a correction — quote this product id: baseten.