Public Endpoints
Instant API access to pre-deployed third-party AI models for image, video, audio and text generation, billed per request or per token with no infrastructure setup.
Find alternatives to Public Endpoints
Per-request / per-token, varying by model: 'Flux Dev costs $0.02 per megapixel'; 'text generation via Qwen3 32B runs $10.00 per 1M tokens'. Failed generations are not charged.
What it does
- Model inference
- Model hosting
- Image generation
- Video generation
- Text to speech
- Text generation
Sources
- Pricing
-
Fetched 2026-09-06; page content returned. Numeric HTTP status unavailable this run — see the batch's method note. Public Endpoints section lists per-model audio/image/language/video rates on a per-request or per-token basis; docs quote 'Flux Dev... $0.02 per megapixel' and 'Qwen3 32B... $10.00 per 1M tokens'.
- Description
-
Official documentation · 6 Sep 2026
Fetched 2026-09-06; page content returned. Numeric HTTP status unavailable this run — see the batch's method note. Docs state: 'Runpod offers Public Endpoints for instant API access to pre-deployed AI models for image, video, audio, and text generation.' Admitted under the docs-only rule (batch 8): the docs give it a name and the pricing page gives it a price line.
- Sold within
-
runpod.io/pricing (already cited pricing-page elsewhere on this record); own top-level pricing entry.
- Interfaces
-
Official documentation · 29 Sep 2026
Docs offer 'the playground and REST API' for requests, and state 'the @runpod/ai-sdk-provider package integrates Public Endpoints with the Vercel AI SDK, providing a streamlined interface for text generation, streaming, and image generation.'
Also from RunPod
-
Pods Infrastructure service
Per-hour GPU pods and per-hour serverless endpoints across both datacentre accelerators and consumer cards, sold on price — the company's own claim is compute up to 90% below traditional cloud providers.
-
Serverless Infrastructure service
Autoscaling GPU API endpoints for AI inference, billed per second with scale-to-zero and sub-200ms cold starts.
-
Runpod Clusters Infrastructure service
Multi-node GPU environments with high-speed InfiniBand interconnect for distributed training and large batch workloads.
-
Runpod Hub Developer tool
A catalog of templates, models and open-source AI apps that can be forked and deployed onto Runpod Serverless in one click.
-
Runpod Hybrid Cloud Platform
Brings customer-owned or rented GPU hardware under Runpod's console, CLI and APIs as a single control plane, with Runpod cloud used for overflow capacity.
Something wrong here? Send a correction — quote this product id: runpod-public-endpoints.