Modal Inference

Infrastructure service

Serve, scale and optimise model inference on Modal's runtime, with sub-second cold starts and autoscaling across regions.

Find alternatives to Modal Inference

TypeInfrastructure service
RoleBundled component
AvailabilitySold
DeploymentSaaS
Intended forDeveloperSelf-reported · 8 Sep 2026
SecurityNot recorded
StatusActive
Sold withinModal — Included with the parentOfficial documentation · 5 Sep 2026
PricingFreemium, Usage-basedPricing page · 5 Sep 2026
Sources1 of 1 field

Metered per second of compute on Modal's one rate card — Nvidia B200 $0.001736/sec, H200 SXM $0.001261/sec, H100 SXM5 $0.001097/sec, A100 80GB $0.000694/sec — with $30 a month of free compute. The products are not priced separately.

What it does

modal.com

How it compares

Sources

Pricing

Pricing page · 5 Sep 2026

Metered per second of compute on Modal's one rate card — Nvidia B200 $0.001736/sec, H200 SXM $0.001261/sec, H100 SXM5 $0.001097/sec, A100 80GB $0.000694/sec — with $30 a month of free compute. The pro

Description

Official documentation · 5 Sep 2026

Modal's own product page for it, fetched 2026-09-05, HTTP 200.

Sold within

Official documentation · 5 Sep 2026

Modal's pricing page prices compute, not products: every one of the six entries in its Products nav runs on the same account and the same meter.

Also from Modal

Open Modal in the directory

Something wrong here? Send a correction — quote this product id: modal-inference.