Crusoe Managed Inference

API service

Hosted endpoints serving open-weight and customer-supplied models on Crusoe's own hardware, billed per token or as dedicated hourly deployments.

Find alternatives to Crusoe Managed Inference

TypeAPI service
RoleBundled component
AvailabilitySold
DeploymentSaaS
Intended forDeveloperSelf-reported · 8 Sep 2026
SecurityNot recorded
StatusActive
Sold withinCrusoe Intelligence Foundry — Included with the parentOfficial documentation · 6 Sep 2026
PricingUsage-based, Enterprise quotePricing page · 21 Aug 2026
Sources1 of 1 field

Serverless per-1M-token rates vary by model with separate input, output and cached-token tiers (e.g. Llama 3.3 70B $0.25 input, DeepSeek V3 $0.50 input, GLM 5.2 $1.40 input). Self-Serve Deployments are billed hourly: NVIDIA H100 80GB HGX $5.50/hr, H200 141GB HGX $6.00/hr. Provisioned throughput is transacted via AI Model Units (AMUs) on commitment - contact sales.

What it does

crusoe.ai

Sources

Pricing

Pricing page · 21 Aug 2026

Sold within

Official documentation · 6 Sep 2026

https://www.crusoe.ai/developers - 'Fine-tune and serve models in Crusoe Intelligence Foundry with $5 in free credits.'

Also from Crusoe

Open Crusoe in the directory

Something wrong here? Send a correction — quote this product id: crusoe-managed-inference.