W&B Serverless Inference
Hosted inference service for open-source and commercial LLMs (OpenAI, Qwen, Llama, Kimi, Phi, DeepSeek, Z.AI) without managing infrastructure.
Find alternatives to W&B Serverless Inference
Free credits initially, then $5/month on the Pro plan (wandb.ai/site/pricing).
What it does
- Model inference
- Model hosting
Sources
- Pricing
- Description
- What it does
- Sold within
-
wandb.ai/site/pricing/ states: 'Serverless Inference: Free credits initially; $5/month on Pro' -- an add-on charge layered on the Pro plan, not a standalone entitlement statement.
- Name
- URL
Also from Weights & Biases
-
W&B Models Platform
Experiment tracking for model training runs, recording hyperparameters, metrics and artifacts so runs can be compared, swept and reproduced.
-
W&B Weave Developer tool
Tracing, evaluation and production monitoring for LLM and agent applications, capturing each call so prompts and outputs can be scored over time.
-
W&B Registry Platform
Curated central repository providing versioning, aliases, lineage tracking and governance for models and datasets across the ML lifecycle.
-
W&B Serverless RL Platform
Managed reinforcement-learning fine-tuning service for LLMs on CoreWeave's managed GPU cluster, billed per-token for rollouts with automatic scale-to-zero.
-
W&B Serverless SFT Platform
Serverless supervised fine-tuning for LLMs on CoreWeave's managed GPU cluster, run alongside Serverless RL in a unified workflow via the Agent Reinforcement Trainer (ART) API.
Open Weights & Biases in the directory
Something wrong here? Send a correction — quote this product id: wandb-serverless-inference.