Modal
Serverless GPU compute: a Python decorator puts a function on an accelerator, scales it from zero to thousands of containers and stops billing when it stops running — aimed at inference, fine-tuning and batch jobs rather than reserved clusters.
Starter $0/month plus compute with $30/month of free credits (100 containers, 10 GPU concurrency); Team $250/month plus compute with $100/month credits (5,000 containers, 50 GPU concurrency); Enterprise custom. Compute is billed per SECOND: B300 $0.001972/sec, B200 $0.001736/sec, H200 SXM $0.001261/sec, H100 SXM5 $0.001097/sec, A100 80GB $0.000694/sec, L40S $0.000542/sec, L4 $0.000222/sec, T4 $0.000164/sec, plus CPU at $0.0000131/core/sec and memory at $0.00000222/GiB/sec. Volumes $0.09/GiB/month with 1 TiB free.
What it does
- GPU cloud
- Model inference
- Model hosting
How it compares
-
Baseten
Direct competitors include Baseten, which raised $150M Series D and focuses on mission-critical inference with dedicated deployments
Research report · 19 Sep 2026 -
Together Inference
Specialized AI platforms like Modal, Together.ai, and Fireworks.ai
Official documentation · 19 Sep 2026 -
Fireworks AI
Specialized AI platforms like Modal, Together.ai, and Fireworks.ai
Official documentation · 19 Sep 2026 -
Serverless
Runpod vs Modal: Python-native serverless versus portable GPU compute
Official documentation · 19 Sep 2026
Sources
- Pricing
-
Plan tiers, included credits, per-second GPU/CPU/memory rates and volume pricing all read from the page 2026-09-04.
- Sold within
-
Modal's pricing page offers "Get started" and "Sign up" and lists "Starter: $0 + compute / month", "Team: $250 + compute / month" and "Enterprise: Custom compute", with "$30 / month" of included compute on Starter.
Also from Modal
-
Modal Inference Infrastructure service
Serve, scale and optimise model inference on Modal's runtime, with sub-second cold starts and autoscaling across regions.
-
Modal Training Infrastructure service
Managed training runs on Modal's fleet, configured in Python alongside the rest of a team's code.
-
Modal Sandboxes Infrastructure service
Isolated, instantly-started containers for running untrusted or agent-generated code at scale — the primitive behind AI app-generation products.
-
Modal Notebooks Developer tool
Hosted notebooks backed by Modal's GPUs, for profiling and experimenting without provisioning a machine.
-
Modal Batch Infrastructure service
Batch execution of large jobs across Modal's fleet, described as one line of code on the product page.
Something wrong here? Send a correction — quote this product id: modal-platform.