Two Nvidia data center GPUs, one Grace Blackwell and one Ada Lovelace: 186 GB HBM3e against 48 GB GDDR6, 9.3x the memory bandwidth on the GB200, and NVLink 1,800 GB/s on the GB200 where the L40 is PCIe only.

At a glance

Memory

GB200
186 GB
L40
48 GB

3.9x more on GB200

Memory bandwidth

GB200
8,000 GB/s
L40
864 GB/s

9.3x higher on GB200

Cheapest on-demand

GB200
$10.50 /GPU/hr
L40
$0.79 /GPU/hr

13x more expensive on GB200

Price comparison

Median price per GPU per hour by billing type, with the cheapest listing under it.

Billing GB200 L40 Difference
On-demand $16.00 /GPU/hr from $10.50 $1.05 /GPU/hr from $0.79 15x more expensive on GB200
Reserved -- /GPU/hr from $10.58 -- /GPU/hr from $0.70
Spot -- /GPU/hr from $27.04 $0.80 /GPU/hr from $0.78

Heads up: A provider's own page may quote a different figure, for example on tax, region or a promotion. A monthly-only plan shows a derived hourly rate. Medians are across providers, one vote each, so a large catalog does not outweigh a small one. No median for reserved and spot on the GB200 and reserved on the L40: we only show one when at least 3 providers list the GPU on that billing type. Verify before provisioning. More on how we price.

Price history

GB200 vs L40 price history

Weekly median on-demand price per GPU per hour

Last 12 months
Aggregating historical prices...

Specs comparison

Spec GB200 L40 Difference
Memory
Capacity 186 GB HBM3e 48 GB GDDR6 3.9x more on GB200
Bandwidth 8,000 GB/s 864 GB/s 9.3x higher on GB200
Bus width 384-bit
Compute · dense
FP4 10,000 TFLOPS Not supported GB200 only
FP8 5,000 TFLOPS 362 TFLOPS 14x higher on GB200
FP16 / BF16 2,500 TFLOPS 181.1 TFLOPS 14x higher on GB200
INT8 5,000 TOPS 362 TOPS 14x higher on GB200
FP32 80 TFLOPS 90.5 TFLOPS 1.1x higher on L40
FP64 40 TFLOPS Not supported GB200 only
Platform
Interconnect NVLink 1,800 GB/s PCIe only NVLink on GB200 only
Board power 1,200 W 300 W 4x higher on GB200
Released Q1 2024 Q4 2022 17 months newer on GB200
Precision
Supported formats FP4FP6FP8FP16BF16TF32FP32FP64INT8 FP8FP16BF16TF32FP32INT4INT8

Compute: Vendor peak figures, dense and per GPU. Not measured throughput.

Source: GB200 datasheet and L40 datasheet.

Where to rent them

Providers listing these GPUs, those with both first, so you can switch without moving clouds.

Provider GB200 /GPU/hr L40 /GPU/hr
CoreWeave logo CoreWeave On-Demand $10.50 On-Demand $1.25
Oracle Cloud logo Oracle Cloud On-Demand $16.00 --
Microsoft Azure logo Azure On-Demand $27.04 --
Amazon Web Services logo AWS Reservation $10.58 --
CUDO Compute logo CUDO Custom on request --
Canopy Wave logo Canopy Wave Custom on request --
Nebius logo Nebius Custom on request --
Together AI logo Together AI Custom on request --
Thunder Compute logo Thunder Compute -- On-Demand $0.79
Runcrate logo Runcrate -- On-Demand $0.97
Oblivus logo Oblivus -- On-Demand $1.05
EmpirioLabs AI logo EmpirioLabs AI -- On-Demand $1.50
Spheron logo Spheron -- Spot $0.88
Hyperstack logo Hyperstack -- On-Demand $1.00
TensorDock logo TensorDock -- On-Demand $1.06

8 providers list the GB200 and 12 the L40. Every listing is on the GB200 page and the L40 page.

Similar GPUs

Other comparisons in the same class.

Our data for Nvidia GB200 was last updated on Sept. 7, 2026, and for Nvidia L40 on Sept. 7, 2026.