Together GPU Clusters
Reserved NVIDIA GPU clusters for training and large-scale inference.
Find alternatives to Together GPU Clusters
What it does
- GPU cloud
- Model training
How it compares
-
CoreWeave AI Cloud
CoreWeave is a cloud infrastructure platform optimized for GPU-based workloads.
Research report · 19 Sep 2026 -
Lambda Cloud
Lambda, founded in 2012, is a cloud-based GPU company that specializes in GPU workstations, cloud services, and AI infrastructure.
Research report · 19 Sep 2026 -
Pods
RunPod, founded in 2022, is a cloud-based infrastructure service that provides cost-effective and scalable GPU resources.
Research report · 19 Sep 2026
Sources
- Pricing
- Sold within
-
Together's GPU Clusters page says "Both options are fully self-serve." and lists on-demand rates such as NVIDIA H100 at "$3.99" "/hr per GPU" with a "Create cluster" button; HGX B300 and NVL72 systems are "Contact us for pricing".
Also from Together AI
-
Together Inference Model API
Hosted API serving open-weight text, image and audio models, billed per token.
-
Together Fine-Tuning API service
A managed service for fine-tuning open-source models on a customer's own data and serving the result on Together's infrastructure.
-
Together Batch Inference API service
Asynchronous bulk inference for workloads that do not need a real-time response, priced below Together's serverless rate.
-
Together Custom Training Developer tool
Custom model training service covering supervised fine-tuning and direct preference optimization, billed per token.
-
Together Dedicated Container Inference Infrastructure service
Dedicated, reserved GPU containers for model inference with guaranteed performance, billed per GPU-hour.
Open Together AI in the directory
Something wrong here? Send a correction — quote this product id: together-gpu-clusters.