The L40 is newer, with 3x the memory and 4.3x the bandwidth. The A16 is cheaper right now.
At a glance
Memory
3x more on L40
- A16
- 16 GB
- L40
- 48 GB
Memory bandwidth
4.3x higher on L40
- A16
- 200 GB/s
- L40
- 864 GB/s
Cheapest on-demand
76% cheaper on A16
- A16
- $0.11 /GPU/hr
- L40
- $0.46 /GPU/hr
Price comparison
Median price per GPU per hour by billing type, with the cheapest listing under it.
| Billing | A16 | L40 | Difference |
|---|---|---|---|
| On-demand | -- /GPU/hr from $0.11 | $1.00 /GPU/hr from $0.46 | |
| Reserved | No reserved listings | -- /GPU/hr from $0.70 | |
| Spot | -- /GPU/hr from $0.07 | $0.78 /GPU/hr from $0.40 |
Heads up: A provider's own page may quote a different figure, for example on tax, region or a promotion. A monthly-only plan shows a derived hourly rate. Medians are across providers, one vote each, so a large catalog does not outweigh a small one. No median for on-demand and spot on the A16 and reserved on the L40: we only show one when at least 3 providers list the GPU on that billing type. Verify before provisioning. More on how we price.
Price history
A16 vs L40 price history
Weekly median on-demand price per GPU per hour
Specs comparison
| Spec | A16 | L40 | Difference |
|---|---|---|---|
| Memory | |||
| Capacity | 16 GB GDDR6 | 48 GB GDDR6 | 3x more on L40 |
| Bandwidth | 200 GB/s | 864 GB/s | 4.3x higher on L40 |
| Bus width | 128-bit | 384-bit | 3x wider on L40 |
| Compute · dense | |||
| FP8 | Not supported | 362 TFLOPS | L40 only |
| FP16 / BF16 | 17.9 TFLOPS | 181.1 TFLOPS | 10x higher on L40 |
| INT8 | 35.9 TOPS | 362 TOPS | 10x higher on L40 |
| FP32 | 4.5 TFLOPS | 90.5 TFLOPS | 20x higher on L40 |
| Platform | |||
| Interconnect | PCIe only | PCIe only | |
| Released | Q2 2021 | Q4 2022 | 2 years newer on L40 |
| Precision | |||
| Supported formats | FP16BF16TF32FP32INT4INT8 | FP8FP16BF16TF32FP32INT4INT8 | |
Compute: Vendor peak figures, dense and per GPU. Not measured throughput.
Source: A16 datasheet and L40 datasheet.
Where to rent them
Providers that rent these GPUs, with each one's cheapest rate. Providers that rent both are listed first.
| Provider | A16 | L40 |
|---|---|---|
|
|
$0.11 /GPU/hr On-demand | $0.46 /GPU/hr On-demand |
|
|
$0.56 /GPU/hr On-demand | $0.97 /GPU/hr On-demand |
|
|
$0.47 /GPU/hr On-demand | -- |
|
|
-- | $0.69 /GPU/hr On-demand |
|
|
-- | $0.69 /GPU/hr On-demand |
|
|
-- | $0.79 /GPU/hr On-demand |
|
|
-- | $0.86 /GPU/hr On-demand |
|
|
-- | $1.00 /GPU/hr On-demand |
|
|
-- | $1.14 /GPU/hr On-demand PCIe |
3 providers list the A16 and 13 the L40. Every listing is on the A16 page and the L40 page.
Similar GPUs
Other comparisons in the same class.
-
Nvidia L4
24 GB GDDR6 · 50% more bandwidth than the A16
From $0.32 /GPU/hr Compare -
Nvidia T4
16 GB GDDR6 · 50% more bandwidth than the A16
From $0.15 /GPU/hr Compare -
Nvidia A10
24 GB GDDR6 · 3x the bandwidth of the A16
From $1.00 /GPU/hr Compare -
Nvidia A100
80 GB HBM2e · 2.2x the bandwidth of the L40
From $0.37 /GPU/hr Compare -
Nvidia L40S
48 GB GDDR6 · released Q3 2023
From $0.55 /GPU/hr Compare
Our data for Nvidia A16 was last updated on Sept. 14, 2026, and for Nvidia L40 on Sept. 14, 2026.