Serverless

Infrastructure service

Autoscaling GPU API endpoints for AI inference, billed per second with scale-to-zero and sub-200ms cold starts.

Find alternatives to Serverless

TypeInfrastructure service
RoleSuite
AvailabilitySold
DeploymentSaaS
Intended forDeveloper, EnterpriseSelf-reported · 8 Sep 2026
SecurityNot recorded
StatusActive
Sold withinSold on its ownPricing page · 15 Sep 2026
PricingUsage-basedPricing page · 6 Sep 2026
Sources1 of 1 field

Per-second GPU billing; published tiers run from $0.58/hr (A4000/A4500/RTX 4000/RTX 2000 class) to $9.98/hr (B300). H100 $4.79/hr, A100 $2.72/hr.

What it does

runpod.io

How it compares

Sources

Pricing

Pricing page · 6 Sep 2026

Fetched 2026-09-06; page content returned. Numeric HTTP status unavailable this run — see the batch's method note. Serverless table quotes 'B300: $9.98/hr', 'H100: $4.79/hr', 'A100: $2.72/hr', 'A4000/A4500/RTX 4000/RTX 2000: $0.58/hr'; billing 'per second, metered from worker start to full stop'.

Description

Official documentation · 6 Sep 2026

Fetched 2026-09-06; page content returned. Numeric HTTP status unavailable this run — see the batch's method note. Headline: 'Dedicated Serverless GPU API endpoints'; body: 'Runpod Serverless runs AI inference with sub-200ms FlashBoot cold starts, per-second billing, and scale to zero.'

Sold within

Pricing page · 15 Sep 2026

runpod.io/pricing (already cited pricing-page elsewhere on this record); own top-level pricing entry.

Also from RunPod

Open RunPod in the directory

Something wrong here? Send a correction — quote this product id: runpod-serverless.