Fireworks Inference

API service

Serving for frontier open models and for a customer's own post-trained versions of them, on an inference engine tuned at each layer.

Find alternatives to Fireworks Inference

TypeAPI service
RoleStandalone tool
AvailabilitySold
DeploymentSaaS
Intended forDeveloper, EnterpriseSelf-reported · 8 Sep 2026
SecurityNot recorded
StatusActive
Sold withinSold on its ownPricing page · 6 Sep 2026
PricingUsage-basedPricing page · 6 Sep 2026
Sources1 of 1 field

Metered per token; per-model rates are on the pricing page.

What it does

fireworks.ai

Sources

Pricing

Pricing page · 6 Sep 2026

Fireworks publishes per-token and per-GPU-hour rates; each surface is metered separately.

Description

Official documentation · 6 Sep 2026

Summarised from the page's own title, <h1> and meta description, fetched 2026-09-06.

Sold within

Pricing page · 6 Sep 2026

Fireworks publishes per-token and per-GPU-hour rates; each surface is metered separately.

URL

Official documentation · 6 Sep 2026

Fetched 2026-09-06: HTTP 200, <title> 'Inference | Fireworks', <h1> 'The highest performance inference for your specialized intelligence.'.

Also from Fireworks AI

Open Fireworks AI in the directory

Something wrong here? Send a correction — quote this product id: fireworks-inference.