Nebius
Netherlands 1 config 1x
Grace CPU + Blackwell Ultra GPU superchip built for rack-scale AI reasoning and inference.
The GB300 is listed by 4 cloud providers, but none offer public on-demand pricing.
Heads up: We do our best to keep these specs & prices accurate. However, cloud costs may fluctuate based on region, usage, and other factors not listed here. These are estimates based on common setups and are for informational purposes only. Always verify current rates & exact specs with the provider before provisioning.
Pairs Grace CPUs with Blackwell Ultra GPUs carrying 288GB HBM3e each, connected in a 72-GPU NVLink domain. Suited for frontier-scale training and long-context reasoning inference where per-GPU memory and interconnect bandwidth are the bottleneck.
Rack-scale form factor with limited cloud availability at the highest cost tier. The GB200 offers a similar architecture with wider availability, and the standalone B200 is sufficient for most Blackwell workloads.
The median on-demand price has held steady around $15.36/hr per GPU across providers.
With 288GB of VRAM, the GB300 can usually run 70B-class models with headroom, and may handle much larger models in 4-bit quantized form depending on runtime overhead, KV cache, and context length.
The GB300 has 288GB of VRAM. Multi-GPU setups increase total memory, but that memory is not automatically pooled across GPUs.
The GB300 has 8,000 GB/s of memory bandwidth. Higher bandwidth helps with faster data transfer between GPU memory and compute cores.
The GB300 supports 9 precision formats. Training: BF16, FP16, TF32, FP32. Inference: FP4, FP6, FP8, INT8. Scientific: FP64.
Yes. The GB300 supports NVLink with 1800 GB/s of bidirectional bandwidth. This helps accelerate multi-GPU communication.
The GB300 is available from 4 cloud providers, including Verda, Amazon Web Services, Gcore. Pricing and availability vary by region and billing model.
Yes. We currently track 24 GB300 listings across 4 cloud providers:
| Billing type | Listings | Avg $/GPU/hr |
|---|---|---|
| On-demand | 3 | $8.62/hr |
| Reserved | 16 | $7.90/hr |
| Spot | 3 | $4.31/hr |
| Custom contract | 2 | Custom |
| Configuration | 36 Grace CPUs : 72 Blackwell Ultra GPUs |
| FP4 Tensor Core | 1,440 PFLOPS¹ | 1,080 PFLOPS |
| FP8/FP6 Tensor Core¹ | 720 PFLOPS |
| INT8 Tensor Core¹ | 24 POPS |
| FP16/BF16 Tensor Core¹ | 360 PFLOPS |
| TF32 Tensor Core¹ | 180 PFLOPS |
| FP32 | 6 PFLOPS |
| FP64 | 100 TFLOPS |
| GPU Memory | Bandwidth | Up to 20 TB HBM3e | 576 TB/s |
| NVLink Bandwidth | 130 TB/s |
| CPU Core Count | 2,592 Arm® Neoverse V2 cores |
| CPU Memory | Bandwidth | Up to 17 TB LPDDR5X | Up to 14.3 TB/s |
¹ With sparsity.
Source: official Nvidia GB300 datasheet.
Previous Grace Blackwell superchip with 192GB GPUs. From $10.50/hr per GPU across 7 providers.
Standalone Blackwell Ultra GPU in x86 systems. From $2.99/hr per GPU across 25 providers.
Standalone Blackwell GPU with wider availability. From $3.35/hr per GPU across 27 providers.
Last updated