Nvidia GH200
Grace Hopper superchip combining ARM CPU with H100 GPU.
- Arch
- Grace Hopper
- Memory
- 96 GB HBM3
- Bandwidth
- 4.00 TB/s
- Released
- Q2 2023
Price history
Weekly median price per GPU per hour
At a glance Updated 9 minutes ago
- Cheapest
- $6.50 /GPU/hr
- on-demand · not verified in stock
- Median price (current)
- --
- 90-day trend
- --
- Coverage
- 4 providers
- 5 price points
Pricing & availability
How this list works
The two views
By provider shows one card per company, placed where its best offer ranked. By configuration lists every offer.
Order
What you can rent comes first: in stock, then waitlist, then not reported, then out of stock. Priced offers rank ahead of quote-only, and on-demand ahead of other billing types.
Within a group, five factors set the order:
- Location: datacenter proximity, blended with provider HQ.
- Price: hourly price, per GPU and in total.
- Billing type: reserved ahead of spot, spot ahead of quote-only.
- Specs: more VRAM, vCPUs and RAM.
- Provider diversity: a provider's repeat rows rank slightly lower, so one company can't take all the top spots.
Search
We match the provider, its country, GPU model, form factor, billing type, availability and instance name. Matching is partial and case-insensitive. Everyday words work too, so "interruptible" finds spot and "sold out" finds out of stock.
Transparency and funding
- Ads and sponsors: Paid placements sit at the top of the list and are always labeled as sponsored content. Sponsorship never influences the organic ranking itself.
- Affiliates: Affiliate links are marked. We may earn a commission if you click them, but commissions never affect the order.
- Prices: Shown in USD, converted at daily reference rates where a provider publishes in another currency. A month means 720 hours. How we estimate costs.
Heads up: A provider's own page may quote a different figure, for example on tax, region or a promotion. A monthly-only plan shows a derived hourly rate. Verify before provisioning. More on how we price.
What can the GH200 run?
One GH200 has 96 GB of VRAM. In practice, that's enough memory for roughly 146B parameters at 4-bit or 38B at 16-bit, assuming a 32K context. Below are some open-weight LLMs, with the estimated memory and GPUs each one needs.
| Model | Memory (INT4 / FP4) | GH200s needed | Cost /hr | Cost /mo |
|---|---|---|---|---|
|
|
18 GB
|
1
|
–
|
–
|
|
|
22 GB
|
1
|
–
|
–
|
|
|
67 GB
|
1
|
–
|
–
|
|
|
178 GB
|
3
|
–
|
–
|
|
|
239 GB
|
3
|
–
|
–
|
|
|
306 GB
|
4
|
–
|
–
|
|
|
1,544 GB
|
18
|
–
|
–
|
Estimates based on the median on-demand rate. Memory is weights plus FP8 KV cache at 32K context per request (FP16/BF16 in the 16-bit column). GPU counts assume 90% of advertised VRAM is usable. No guarantee of runtime support, usable performance, or that a matching quantized build exists. How we estimate costs.
Technical Specifications
Nvidia GH200 · Per GPU
| Compute · dense | |
|---|---|
| FP8 | 1,979 TFLOPS 3,958 with sparsity |
| FP16 / BF16 | 989.5 TFLOPS 1,979 with sparsity |
| INT8 | 1,979 TOPS 3,958 with sparsity |
| FP32 | 67 TFLOPS |
| FP64 | 34 TFLOPS |
| Precision support | FP8FP16BF16TF32FP32FP64INT8 |
| Memory | |
|---|---|
| Capacity | 96 GB HBM3 |
| Bandwidth | 4,000 GB/s |
| ECC | Yes |
| Silicon | |
|---|---|
| Architecture | Hopper |
| Process | TSMC 4N |
| Transistors | 80 billion |
| Shader cores | 16,896 CUDA cores |
| Matrix cores | 528 Tensor cores |
| Compute units | 132 SMs |
| Fabric and host | |
|---|---|
| GPU interconnect | NVLink 900 GB/s |
Source: official Nvidia GH200 datasheet.
Frequently Asked Questions
How much does the GH200 cost per hour?
-
As of September 13, 2026, we track 5 configs from 4 providers. Prices are per GPU per hour.
Billing type Configs Median /GPU/hr Cheapest On-demand
4
$6.50
Spot
1
Sold out
No median for on-demand: we only show this when at least 3 providers list the GPU on that billing type.
Who has the cheapest GH200?
-
The lowest listed on-demand price for the GH200 is $6.50 per GPU per hour from CoreWeave, though we haven't verified current stock.
Where can I rent a GH200?
-
GetDeploying currently tracks GH200 configs from 4 providers. See the full price comparison above for every provider and config.
How many GH200 GPUs do I need?
-
One GH200 runs models up to roughly 146B parameters at 4-bit quantization or 38B at 16-bit, assuming a 32K context. Larger models run across multiple GPUs: the model table above shows the estimated memory and GPU count for popular open-weight LLMs.
Why choose the GH200?
-
Grace Hopper superchip combines ARM CPU with H100 GPU and 96GB or 144GB HBM3e. Unified memory architecture with 900 GB/s NVLink. Efficient for AI workloads that benefit from tight CPU-GPU integration.
When is the GH200 not a good fit?
-
ARM CPU ecosystem may have compatibility issues with some software stacks. For standard GPU cloud workloads, a regular H100 is simpler to deploy and more widely available.
Alternatives to Nvidia GH200
Last updated