Nvidia GH200

Nvidia GH200

Grace Hopper superchip combining ARM CPU with H100 GPU.

Arch
Grace Hopper
Memory
96 GB HBM3
Bandwidth
4.00 TB/s
Released
Q2 2023

Price history

Weekly median price per GPU per hour

Last 12 months
Aggregating historical prices...

At a glance Updated 9 minutes ago

Cheapest
$6.50 /GPU/hr
on-demand · not verified in stock
Median price (current)
--
90-day trend
--
Coverage
4 providers
5 price points

Pricing & availability

How this list works

The two views

By provider shows one card per company, placed where its best offer ranked. By configuration lists every offer.

Order

What you can rent comes first: in stock, then waitlist, then not reported, then out of stock. Priced offers rank ahead of quote-only, and on-demand ahead of other billing types.

Within a group, five factors set the order:

  1. Location: datacenter proximity, blended with provider HQ.
  2. Price: hourly price, per GPU and in total.
  3. Billing type: reserved ahead of spot, spot ahead of quote-only.
  4. Specs: more VRAM, vCPUs and RAM.
  5. Provider diversity: a provider's repeat rows rank slightly lower, so one company can't take all the top spots.

Search

We match the provider, its country, GPU model, form factor, billing type, availability and instance name. Matching is partial and case-insensitive. Everyday words work too, so "interruptible" finds spot and "sold out" finds out of stock.

Transparency and funding

  • Ads and sponsors: Paid placements sit at the top of the list and are always labeled as sponsored content. Sponsorship never influences the organic ranking itself.
  • Affiliates: Affiliate links are marked. We may earn a commission if you click them, but commissions never affect the order.
  • Prices: Shown in USD, converted at daily reference rates where a provider publishes in another currency. A month means 720 hours. How we estimate costs.
CoreWeave logo

United States of America flag USA 1 config 1x

On-Demand from $6.50
From $6.50 /GPU/hr On-Demand Visit website
Lambda Labs logo

Lambda Labs

Out of stock

United States of America flag USA 1 config 1x

On-Demand from $2.29
From $2.29 /GPU/hr On-Demand Visit website
Sesterce logo

Sesterce

Out of stock

France flag France 2 configs 1x

On-Demand from $2.52
From $2.52 /GPU/hr On-Demand Visit website
Vultr logo

Vultr

Out of stock

United States of America flag USA 1 config 1x

Spot from $1.99
From $1.99 /GPU/hr Spot Visit website

Heads up: A provider's own page may quote a different figure, for example on tax, region or a promotion. A monthly-only plan shows a derived hourly rate. Verify before provisioning. More on how we price.

What can the GH200 run?

One GH200 has 96 GB of VRAM. In practice, that's enough memory for roughly 146B parameters at 4-bit or 38B at 16-bit, assuming a 32K context. Below are some open-weight LLMs, with the estimated memory and GPUs each one needs.

Fits on 1 Multi-node
Model Memory (INT4 / FP4) GH200s needed Cost /hr Cost /mo
Alibaba Cloud logo Qwen3.8-27B 27B
18 GB
1
Google Cloud logo Gemma 4 31B 30.7B
22 GB
1
OpenAI logo GPT-OSS-120B 117B · 5.1B active
67 GB
1
Z.AI logo GLM-5.3-Flash 320B · 18B active
178 GB
3
MiniMax logo MiniMax-M3 428B · 23B active
239 GB
3
DeepSeek logo DeepSeek V4.1 Flash 552B · 16B active
306 GB
4
Moonshot AI logo Kimi K3 2.8T · 104B active
1,544 GB
18

Estimates based on the median on-demand rate. Memory is weights plus FP8 KV cache at 32K context per request (FP16/BF16 in the 16-bit column). GPU counts assume 90% of advertised VRAM is usable. No guarantee of runtime support, usable performance, or that a matching quantized build exists. How we estimate costs.

Technical Specifications

Nvidia GH200 · Per GPU

Compute · dense
FP8 1,979 TFLOPS 3,958 with sparsity
FP16 / BF16 989.5 TFLOPS 1,979 with sparsity
INT8 1,979 TOPS 3,958 with sparsity
FP32 67 TFLOPS
FP64 34 TFLOPS
Precision support FP8FP16BF16TF32FP32FP64INT8
Memory
Capacity 96 GB HBM3
Bandwidth 4,000 GB/s
ECC Yes
Silicon
Architecture Hopper
Process TSMC 4N
Transistors 80 billion
Shader cores 16,896 CUDA cores
Matrix cores 528 Tensor cores
Compute units 132 SMs
Fabric and host
GPU interconnect NVLink 900 GB/s

Source: official Nvidia GH200 datasheet.

Frequently Asked Questions

How much does the GH200 cost per hour?

As of September 13, 2026, we track 5 configs from 4 providers. Prices are per GPU per hour.

Billing typeConfigsMedian /GPU/hrCheapest

On-demand

4

$6.50

Spot

1

Sold out

No median for on-demand: we only show this when at least 3 providers list the GPU on that billing type.

Who has the cheapest GH200?

The lowest listed on-demand price for the GH200 is $6.50 per GPU per hour from CoreWeave, though we haven't verified current stock.

Where can I rent a GH200?

GetDeploying currently tracks GH200 configs from 4 providers. See the full price comparison above for every provider and config.

How many GH200 GPUs do I need?

One GH200 runs models up to roughly 146B parameters at 4-bit quantization or 38B at 16-bit, assuming a 32K context. Larger models run across multiple GPUs: the model table above shows the estimated memory and GPU count for popular open-weight LLMs.

Why choose the GH200?

Grace Hopper superchip combines ARM CPU with H100 GPU and 96GB or 144GB HBM3e. Unified memory architecture with 900 GB/s NVLink. Efficient for AI workloads that benefit from tight CPU-GPU integration.

When is the GH200 not a good fit?

ARM CPU ecosystem may have compatibility issues with some software stacks. For standard GPU cloud workloads, a regular H100 is simpler to deploy and more widely available.

Alternatives to Nvidia GH200

Last updated