CoreWeave In stock
USA 1 config 4x
Grace CPU + Blackwell GPU superchip with unified memory architecture.
We're tracking 8 cloud providers for the GB200, and the pricing spread is significant. While the average sits at $17.85/hr, the lowest price is currently $10.50/hr per GPU (on-demand).
USA 1 config 4x
USA 2 configs 4x
USA 1 config 4x
USA 1 config 72x
USA 1 config 72x
USA 1 config 4x
UK 1 config 1x
Netherlands 1 config 1x
Heads up: We do our best to keep these specs & prices accurate. However, cloud costs may fluctuate based on region, usage, and other factors not listed here. These are estimates based on common setups and are for informational purposes only. Always verify current rates & exact specs with the provider before provisioning.
Combined Grace CPU + Blackwell GPU with 192GB HBM3e and 1800 GB/s NVLink. Unified memory architecture eliminates CPU-GPU bottlenecks for large-scale AI training.
Specialized superchip form factor with limited cloud availability. The standalone B200 is more widely offered and sufficient for most Blackwell workloads.
The median on-demand price across providers has risen about 42% since August 2025, from $13.25 to $18.77/hr per GPU. This reflects the market-wide median, which moves both when providers change prices and when lower or higher priced offerings enter the market.
With 192GB of VRAM, the GB200 can usually run 70B-class models with headroom, and may handle much larger models in 4-bit quantized form depending on runtime overhead, KV cache, and context length.
The GB200 has 192GB of VRAM. Multi-GPU setups increase total memory, but that memory is not automatically pooled across GPUs.
The GB200 has 8,000 GB/s of memory bandwidth. Higher bandwidth helps with faster data transfer between GPU memory and compute cores.
The GB200 supports 9 precision formats. Training: BF16, FP16, TF32, FP32. Inference: FP4, FP6, FP8, INT8. Scientific: FP64.
Yes. The GB200 supports NVLink with 1800 GB/s of bidirectional bandwidth. This helps accelerate multi-GPU communication.
GB200 pricing currently ranges from $10.50/hr to $27.04/hr per GPU, depending on the provider, instance type, and billing model.
At 720 hours per month, one GB200 can cost between $7,560.00 to $19,468.80 per month, depending on the provider. Reserved and spot pricing can lower that further.
The GB200 is available from 8 cloud providers, including Amazon Web Services, CoreWeave, Nebius. Pricing and availability vary by region and billing model.
Yes. We currently track 9 GB200 listings across 8 cloud providers:
| Billing type | Listings | Avg $/GPU/hr |
|---|---|---|
| On-demand | 3 | $17.85/hr |
| Reserved | 1 | $10.58/hr |
| Spot | 1 | $27.04/hr |
| Custom contract | 4 | Custom |
| Configuration | 36 Grace CPU : 72 Blackwell GPUs |
| FP4 Tensor Core¹ | 1,440 PFLOPS |
| FP8/FP6 Tensor Core¹ | 720 PFLOPS |
| INT8 Tensor Core¹ | 720 POPS |
| FP16/BF16 Tensor Core¹ | 360 PFLOPS |
| TF32 Tensor Core | 180 PFLOPS |
| FP32 | 5,760 TFLOPS |
| FP64 | 2,880 TFLOPS |
| FP64 Tensor Core | 2,880 TFLOPS |
| GPU Memory | Bandwidth | Up to 13.4 TB HBM3e | 576 TB/s |
| NVLink Bandwidth | 130TB/s |
| CPU Core Count | 2,592 Arm® Neoverse V2 cores |
| CPU Memory | Bandwidth | Up to 17 TB LPDDR5X | Up to 18.4 TB/s |
¹ With sparsity.
Source: official Nvidia GB200 datasheet.
Blackwell Ultra successor with 288GB GPUs. From $3.13/hr per GPU across 7 providers.
Standalone Blackwell GPU with wider availability. From $3.35/hr per GPU across 28 providers.
Previous-gen Grace Hopper superchip. From $1.99/hr per GPU across 4 providers.
Standard Hopper GPU with broad cloud support. From $0.35/hr per GPU across 52 providers.
Last updated