Nvidia H100
The standard data center GPU for large-scale AI training and inference.
- Arch
- Hopper
- Memory
- 80 GB HBM3 SXM
- Bandwidth
- 3.35 TB/s SXM
- Released
- Q3 2022
Price history
Weekly median price per GPU per hour
At a glance Updated 3 minutes ago
- Cheapest in stock
- $1.45 /GPU/hr
- on-demand · 1x SXM · Lium
- Median price (current)
- $3.39 /GPU/hr
- on-demand
- 90-day trend
- +10%
- median on-demand
- Coverage
- 53 providers
- 445 price points
Pricing & availability
How this list works
The two views
By provider shows one card per company, placed where its best offer ranked. By configuration lists every offer.
Order
What you can rent comes first: in stock, then waitlist, then not reported, then out of stock. Priced offers rank ahead of quote-only, and on-demand ahead of other billing types.
Within a group, five factors set the order:
- Location: datacenter proximity, blended with provider HQ.
- Price: hourly price, per GPU and in total.
- Billing type: reserved ahead of spot, spot ahead of quote-only.
- Specs: more VRAM, vCPUs and RAM.
- Provider diversity: a provider's repeat rows rank slightly lower, so one company can't take all the top spots.
Search
We match the provider, its country, GPU model, form factor, billing type, availability and instance name. Matching is partial and case-insensitive. Everyday words work too, so "interruptible" finds spot and "sold out" finds out of stock.
Transparency and funding
- Ads and sponsors: Paid placements sit at the top of the list and are always labeled as sponsored content. Sponsorship never influences the organic ranking itself.
- Affiliates: Affiliate links are marked. We may earn a commission if you click them, but commissions never affect the order.
- Prices: Shown in USD, converted at daily reference rates where a provider publishes in another currency. A month means 720 hours. How we estimate costs.
Runpod
In stockEmpirioLabs AI
In stockRuncrate
In stockMassed Compute
In stockLambda Labs
In stockNovita
In stockThunder Compute
In stockOblivus
In stockKoyeb
In stockSesterce
In stockUpCloud
In stockLyceum
In stockVerda
In stockScaleway
In stockNo providers match .
Heads up: A provider's own page may quote a different figure, for example on tax, region or a promotion. A monthly-only plan shows a derived hourly rate. Verify before provisioning. More on how we price.
What can the H100 run?
One H100 has 80 GB of VRAM on the SXM variant. In practice, that's enough memory for roughly 120B parameters at 4-bit or 31B at 16-bit, assuming a 32K context. Below are some open-weight LLMs, with the estimated memory and GPUs each one needs.
| Model | Memory (INT4 / FP4) | H100s needed | Cost /hr | Cost /mo |
|---|---|---|---|---|
|
|
18 GB
|
1
|
$3.39
|
$2,441
|
|
|
22 GB
|
1
|
$3.39
|
$2,441
|
|
|
67 GB
|
1
|
$3.39
|
$2,441
|
|
|
178 GB
|
4
|
$13.56
|
$9,763
|
|
|
239 GB
|
4
|
$13.56
|
$9,763
|
|
|
306 GB
|
8
|
$27.12
|
$19,526
|
|
|
1,544 GB
|
22
|
–
|
–
|
Estimates based on the median on-demand rate. Memory is weights plus FP8 KV cache at 32K context per request (FP16/BF16 in the 16-bit column). GPU counts assume 90% of advertised VRAM is usable. No guarantee of runtime support, usable performance, or that a matching quantized build exists. How we estimate costs.
Technical Specifications
Nvidia H100 · Per GPU · SXM
| Compute · dense | |
|---|---|
| FP8 | 1,979 TFLOPS 3,958 with sparsity |
| FP16 / BF16 | 989.5 TFLOPS 1,979 with sparsity |
| INT8 | 1,979 TOPS 3,958 with sparsity |
| FP32 | 67 TFLOPS |
| FP64 | 34 TFLOPS |
| Precision support | FP8FP16BF16TF32FP32FP64INT8 |
| Memory | |
|---|---|
| Capacity | 80 GB HBM3 |
| Bandwidth | 3,350 GB/s |
| Bus width | 5,120-bit |
| ECC | Yes |
| Silicon | |
|---|---|
| Architecture | Hopper |
| Process | TSMC 4N |
| Transistors | 80 billion |
| Shader cores | 16,896 CUDA cores |
| Matrix cores | 528 Tensor cores |
| Compute units | 132 SMs |
| Fabric and host | |
|---|---|
| GPU interconnect | NVLink 900 GB/s |
| Host interface | PCIe 5.0 x16 |
| Power | |
|---|---|
| Board power | 700 W |
| Platform | |
|---|---|
| Partitioning | MIG, up to 7 instances |
Source: official Nvidia H100 datasheet.
Frequently Asked Questions
What's the difference between the H100 SXM, PCIe and NVL?
-
The H100 ships in 3 versions with different memory, bandwidth or interconnect. The figures at the top of this page are for the H100 SXM.
Variant Memory Memory bandwidth Interconnect H100 SXM
80 GB
3,350 GB/s
NVLink 900 GB/s
H100 PCIe
80 GB
2,000 GB/s
NVLink 600 GB/s
H100 NVL
94 GB
3,938 GB/s
NVLink 600 GB/s
How much does the H100 cost per hour?
-
As of September 12, 2026, the median on-demand price is $3.39 per GPU per hour across 40 providers with a priced on-demand config.
Billing type Configs Median /GPU/hr Cheapest On-demand
207
$3.39
$1.45 (in stock, Lium)
Reserved
173
$3.10
$1.53 (1 mo, in stock, HyperAI)
Spot
57
$1.94
$0.61 (in stock, Vast.ai)
6 providers also quote the H100 on a custom contract, priced per deal.
Who has the cheapest H100?
How much does the H100 cost per month?
-
At 720 hours per month, one H100 costs an estimated $2,441 at the median on-demand price. Cheapest verified in stock: $1,044 per month on-demand, $1,102 reserved (1 mo), $439 spot.
Where can I rent an H100?
-
GetDeploying currently tracks H100 configs from 53 providers. The cheapest verified in-stock on-demand configs come from Lium, Vast.ai, HyperAI and UpCloud. See the full price comparison above for every provider and config.
Are H100 prices going up or down?
-
As of September 12, 2026, the median on-demand price has risen about 10% over the past 90 days to $3.39 per GPU per hour, and about 11% over the past 12 months.
How many H100 GPUs do I need?
-
One H100 runs models up to roughly 120B parameters at 4-bit quantization or 31B at 16-bit, assuming a 32K context. Larger models run across multiple GPUs: the model table above shows the estimated memory and GPU count for popular open-weight LLMs.
Why choose the H100?
-
80GB HBM3 with 900 GB/s NVLink for multi-node training. FP8 Transformer Engine for improved training throughput over A100. The standard choice for large-scale AI training and high-throughput inference.
Alternatives to Nvidia H100
Last updated