The RTX 5070 is newer and has 3.4x the bandwidth, while the A16 has 33% more memory. The A16 is cheaper right now.
At a glance
Memory
33% more on A16
- A16
- 16 GB
- RTX 5070
- 12 GB
Memory bandwidth
3.4x higher on RTX 5070
- A16
- 200 GB/s
- RTX 5070
- 672 GB/s
Cheapest on-demand
34% cheaper on A16
- A16
- $0.11 /GPU/hr
- RTX 5070
- $0.17 /GPU/hr
Price comparison
Median price per GPU per hour by billing type, with the cheapest listing under it.
| Billing | A16 | RTX 5070 | Difference |
|---|---|---|---|
| On-demand | -- /GPU/hr from $0.11 | $0.19 /GPU/hr from $0.17 | |
| Reserved | No reserved listings | -- /GPU/hr from $0.19 (3 mo) | |
| Spot | -- /GPU/hr from $0.07 | -- /GPU/hr from $0.08 |
Heads up: A provider's own page may quote a different figure, for example on tax, region or a promotion. A monthly-only plan shows a derived hourly rate. Medians are across providers, one vote each, so a large catalog does not outweigh a small one. No median for on-demand and spot on the A16 and reserved and spot on the RTX 5070: we only show one when at least 3 providers list the GPU on that billing type. Verify before provisioning. More on how we price.
Price history
A16 vs RTX 5070 price history
Weekly median on-demand price per GPU per hour
Specs comparison
| Spec | A16 | RTX 5070 | Difference |
|---|---|---|---|
| Memory | |||
| Capacity | 16 GB GDDR6 | 12 GB GDDR7 | 33% more on A16 |
| Bandwidth | 200 GB/s | 672 GB/s | 3.4x higher on RTX 5070 |
| Bus width | 128-bit | 192-bit | 50% wider on RTX 5070 |
| Compute · dense | |||
| FP4 | Not supported | 493.9 TFLOPS | RTX 5070 only |
| FP8 | Not supported | 123.5 TFLOPS | RTX 5070 only |
| FP16 / BF16 | 17.9 TFLOPS | 61.7 TFLOPS | 3.4x higher on RTX 5070 |
| INT8 | 35.9 TOPS | 246.9 TOPS | 6.9x higher on RTX 5070 |
| FP32 | 4.5 TFLOPS | 30.9 TFLOPS | 6.9x higher on RTX 5070 |
| Platform | |||
| Interconnect | PCIe only | PCIe only | |
| Released | Q2 2021 | Q1 2025 | 4 years newer on RTX 5070 |
| Precision | |||
| Supported formats | FP16BF16TF32FP32INT4INT8 | FP4FP8FP16BF16TF32FP32INT8 | |
Compute: Vendor peak figures, dense and per GPU. Not measured throughput.
Source: A16 datasheet and RTX 5070 datasheet.
Where to rent them
Providers that rent these GPUs, with each one's cheapest rate. Providers that rent both are listed first.
| Provider | A16 | RTX 5070 |
|---|---|---|
|
|
$0.11 /GPU/hr On-demand | $0.18 /GPU/hr On-demand |
|
|
$0.47 /GPU/hr On-demand | -- |
|
|
$0.56 /GPU/hr On-demand | -- |
|
|
-- | $0.17 /GPU/hr On-demand |
|
|
-- | $0.23 /GPU/hr On-demand |
Similar GPUs
Other comparisons in the same class.
-
Nvidia L4
24 GB GDDR6 · 50% more bandwidth than the A16
From $0.27 /GPU/hr Compare -
Nvidia T4
16 GB GDDR6 · 50% more bandwidth than the A16
From $0.15 /GPU/hr Compare -
Nvidia A10
24 GB GDDR6 · 3x the bandwidth of the A16
From $0.24 /GPU/hr Compare -
Nvidia RTX 5060 Ti
16 GB GDDR7 · 24% less bandwidth than the RTX 5070
From $0.10 /GPU/hr Compare -
Nvidia RTX 5070 Ti
16 GB GDDR7 · 33% more bandwidth than the RTX 5070
From $0.16 /GPU/hr Compare
Our data for Nvidia A16 was last updated on Sept. 14, 2026, and for Nvidia RTX 5070 on Sept. 14, 2026.