Nvidia logo

A16 vs L4

The A16 packs 4 GPUs for VDI, the L4 is a single GPU for inference. Choose A16 for multi-user virtual desktops, L4 for AI inference.

Overview

Nvidia logo Nvidia A16

The Nvidia A16 is a data center GPU on the Ampere architecture, released Q2 2021, with 16 GB GDDR6 and 200 GB/s memory bandwidth.

Budget Datacenter

Nvidia logo Nvidia L4

The Nvidia L4 is a data center GPU on the Ada Lovelace architecture, released Q1 2023, with 24 GB GDDR6 and 300 GB/s memory bandwidth.

Mid-Range Datacenter

Memory per GPU

1.5x more on L4
Nvidia logo A16
16 GB GDDR6
Nvidia logo L4
24 GB GDDR6

Memory bandwidth

1.5x higher on L4
Nvidia logo A16
200 GB/s
Nvidia logo L4
300 GB/s

GPU interconnect

Nvidia logo A16
PCIe only
Nvidia logo L4
PCIe only

Peak performance (FP16)

6.8x higher on L4
Nvidia logo A16
17.9 TFLOPS
Nvidia logo L4
121 TFLOPS

Vendor peak figures, dense and per GPU. Not measured throughput.

Precision support

Nvidia logo A16
FP16 BF16 TF32 FP32 INT4 INT8
Nvidia logo L4
FP8 FP16 BF16 TF32 FP32 INT4 INT8

Highlighted formats are supported by only one of the two.

Architecture

Nvidia logo A16
Ampere
Nvidia logo L4
Ada Lovelace

Released

2 years newer on L4
Nvidia logo A16
Q2 2021
Nvidia logo L4
Q1 2023

Price comparison

Cheapest on-demand

3.1x higher on L4
Nvidia logo A16
$0.11 / GPU / hr
Nvidia logo L4
$0.33 / GPU / hr

Typical rate

1.8x higher on L4
Nvidia logo A16
$0.52 / GPU / hr
Nvidia logo L4
$0.91 / GPU / hr

Median across providers, one vote each, so a large catalog does not outweigh a small one.

Spot floor

2x higher on L4
Nvidia logo A16
$0.07 / GPU / hr
Nvidia logo L4
$0.13 / GPU / hr

Providers listing it

13 more on L4
Nvidia logo A16
4
Nvidia logo L4
17

Price, last 90 days

Nvidia logo A16
Too few providers pricing it
Nvidia logo L4
+3%

Change in the market-wide median, which moves when providers reprice and when cheaper or more expensive listings enter the market.

Heads up: A provider's own page may quote a different figure, for example on tax, region or a promotion. A monthly-only plan shows a derived hourly rate. Verify before provisioning. More on how we price.

Offers and availability

How this list works

Order

Each provider appears once per GPU. We pick the one configuration that speaks for that provider, and rank on it in this order:

  1. Availability: in stock, then waitlist, then not reported, then out of stock. A price you can act on outranks a cheaper one you cannot.
  2. Billing type: on-demand first, then reserved, then spot, then quote-only. On-demand leads because it is the only rate you can hold without committing to anything.
  3. Price: the lowest per-GPU hourly rate the provider lists, converted to USD.

The configuration we pick is chosen the same way, availability before billing before price. The rest of that provider's configurations are on the GPU's own page.

Ties sort alphabetically. Reserved rates are not separated by term length, so a three-year commitment can sit above a one-month one: the label beneath each price says which you are looking at. Providers who publish no rate are listed last within their group.

Transparency and funding

  • Ads and sponsors: Paid placements sit at the top of the list and are always labeled as sponsored content. Sponsorship never influences the organic ranking itself.
  • Affiliates: Affiliate links are marked. We may earn a commission if you click them, but commissions never affect the order.
  • Prices: Shown in USD, converted at daily reference rates where a provider publishes in another currency. A month means 720 hours. How we estimate costs.

Similar GPUs

Other comparisons in the same class:

View all GPUs

Our data for Nvidia A16 was last updated on Sept. 3, 2026, and for Nvidia L4 on Sept. 3, 2026.