Nvidia H100
80GB 4 configs 1x
GPU cloud for AI model training and inference
Founded in 2020, Together AI is an AI acceleration cloud platform for training, fine-tuning, and running generative AI models. They support over 200 open-source models with both serverless and dedicated GPU clusters.
Together AI stands out with research innovations like FlashAttention-3, offering up to 4x faster inference and 11x lower cost compared to GPT-4o. Their platform features high-performance GPUs including NVIDIA H100, H200, and GB200.
Example customers include Pika Labs, Nexusflow, Arcee AI.
Together Homepage
Together AI uses a per-minute billing model with transparent pricing across all supported GPU models.
Here are some of the GPU configurations offered by Together:
80GB 4 configs 1x
288GB 1 config 72x
192GB 4 configs 1x
141GB 4 configs 1x
192GB 1 config 72x
Heads up: We do our best to keep these specs & prices accurate. However, cloud costs may fluctuate based on region, usage, and other factors not listed here. These are estimates based on common setups and are for informational purposes only. Always verify current rates & exact specs with the provider before provisioning. You can find Together's latest pricing here.
Here are some of the services that Together offers:
Based on our records, Together operates in the following locations:
| Country | Location | Slug |
|---|---|---|
| Canada | Montreal | ca-central-1 |
| Canada | Prince George | ca-west-1 |
| France | Paris | eu-west-2 |
| Germany | Frankfurt | eu-central-1 |
| Iceland | Reykjavik | eu-north-1 |
| Indonesia | Jakarta | asia-southeast-1 |
| Japan | Tokyo | asia-northeast-1 |
| Portugal | Lisbon | eu-west-3 |
| Saudi Arabia | Riyadh | me-central-1 |
| South Korea | Seoul | asia-northeast-2 |
| Spain | Madrid | eu-west-4 |
| UK | London | eu-west-1 |
| USA | Charlotte, North Carolina | us-east-3 |
| USA | Chaska, Minnesota | us-central-1 |
| USA | Chicago, Illinois | us-central-4 |
| USA | Columbus, Ohio | us-central-5 |
| USA | Dallas, Texas | us-south-1 |
| USA | Des Moines, Iowa | us-central-2 |
| USA | Englewood, Colorado | us-west-3 |
| USA | Memphis, Tennessee | us-south-2 |
| USA | Missoula, Montana | us-west-1 |
| USA | Newark, Delaware | us-east-1 |
| USA | Phoenix, Arizona | us-west-4 |
| USA | Salt Lake City, Utah | us-west-2 |
| USA | Silver Spring, Maryland | us-east-2 |
| USA | Sullivan, Missouri | us-central-3 |
Compare Together against other cloud providers:
Runpod offers a wider range of GPU models and flexible deployment options with both dedicated instances and serverless containers.
Salad offers distributed GPU computing using consumer hardware for cost-effective AI training and inference.
Fal provides serverless GPU functions with automatic scaling for AI inference workloads.
Lambda Labs specializes in GPU cloud infrastructure with high-performance instances and private cloud options.
Replicate offers a platform for running pre-built AI models without infrastructure management.
Our data for Together was last updated on July 1, 2026.