Inference Endpoints

Infrastructure service

Deploys a model from the Hub to a dedicated managed endpoint, billed by the hour.

Find alternatives to Inference Endpoints

TypeInfrastructure service
RoleStandalone tool
AvailabilitySold
Deployment—
Intended forDeveloper, EnterpriseSelf-reported · 8 Sep 2026
SecurityNot recorded
StatusActive
Sold withinSold on its ownPricing page · 29 Sep 2026
PricingUsage-based, Enterprise quotePricing page · 21 Aug 2026
Sources1 of 1 field

Free Hub access, PRO at $9 a month, Team at $20 and Enterprise at $50 per user, plus hourly compute for Spaces and Inference Endpoints.

What it does

endpoints.huggingface.co

Sources

Pricing

Pricing page · 21 Aug 2026

Free Hub access, PRO at $9 a month, Team at $20 and Enterprise at $50 per user, plus hourly compute for Spaces and Inference Endpoints.

Sold within

Pricing page · 29 Sep 2026

Its own hourly rate card on the pricing page: "Dedicated inference, starting at $0.033/hour.", followed by per-instance CPU/accelerator/GPU prices. Docs (huggingface.co/docs/inference-endpoints/pricing) add a prerequisite that does not name a plan: "🤗 Inference Endpoints is accessible to Hugging Face accounts with an active subscription and credit card on file."

Also from Hugging Face

Open Hugging Face in the directory

Something wrong here? Send a correction — quote this product id: huggingface-inference-endpoints.