Nvidia A16

Nvidia A16

Purpose-built for virtual desktop (VDI) deployments.

Last update 6 minutes ago
Compare vs other GPUs →
Aggregating historical prices...

Key Specifications

Architecture
Ampere
Memory
4x 16GB GDDR6
Memory Bandwidth
4x 200 GB/s
Release date
Q2 2021

A16 Pricing and Availability

We don't see much volatility for the A16 right now. Most providers are clustered between $0.47 and $0.56/hr per GPU, so availability is likely the deciding factor.

3 providers

How this list works

The two views

By provider shows one card per company, placed where its best offer ranked. By configuration lists every offer.

Order

What you can rent comes first: in stock, then waitlist, then not reported, then out of stock. Priced offers rank ahead of quote-only, and on-demand ahead of other billing types.

Within a group, five factors set the order, largest first:

  1. Location: datacenter proximity, blended with provider HQ.
  2. Price: hourly price, per GPU and in total.
  3. Billing type: reserved ahead of spot, spot ahead of quote-only.
  4. Specs: more VRAM, vCPUs and RAM.
  5. Provider diversity: in By configuration, a provider's repeat rows rank slightly lower, so one company can't take all the top spots.

Search

We match the provider, GPU model, form factor, billing type, availability and instance name. Matching is partial and case-insensitive. Everyday words work too, so "interruptible" finds spot and "sold out" finds out of stock.

Transparency and funding

  • Ads and sponsors: Paid placements sit at the top of the list and are always labeled as sponsored content. Sponsorship never influences the organic ranking itself.
  • Affiliates: Affiliate links are marked. We may earn a commission if you click them, but commissions never affect the order.
  • Prices: Shown in USD, converted at daily reference rates where a provider publishes in another currency. A month means 720 hours. How we estimate costs.
Vultr logo

Vultr In stock

USA 3 configs 2x-8x

On-Demand from $0.47
From $0.47 / GPU / hr On-Demand Visit website
Runcrate logo

Runcrate In stock

USA 1 config 2x

On-Demand from $0.56
From $0.56 / GPU / hr On-Demand PCIe Visit website
Sesterce logo

Sesterce In stock

France 2 configs 2x-4x

On-Demand from $0.56
From $0.56 / GPU / hr On-Demand Visit website

Heads up: We do our best to keep these specs & prices accurate. However, cloud costs may fluctuate based on region, usage, and other factors not listed here. These are estimates based on common setups and are for informational purposes only. Always verify current rates & exact specs with the provider before provisioning.

Frequently Asked Questions

Why choose the A16?

4x 16GB GPUs on a single card (64GB total). Designed for virtual desktop (VDI) and multi-user GPU sharing. Hardware video encode/decode support.

When is the A16 not a good fit?

Built for VDI and multi-user GPU sharing. Individual GPU dies have limited compute, so better suited for remote desktops and lightweight graphics than AI workloads.

Are A16 prices going up or down?

The median on-demand price across providers has risen about 8% since August 2025, from $0.52 to $0.56/hr per GPU. This reflects the market-wide median, which moves both when providers change prices and when lower or higher priced offerings enter the market.

What size AI models can the A16 run?

With 16GB of VRAM, the A16 is best for 7B-class models in 4-bit or 8-bit quantized form, and smaller models in FP16.

How much VRAM does the A16 have?

The A16 has 16GB of VRAM. Multi-GPU setups increase total memory, but that memory is not automatically pooled across GPUs.

What is the A16's memory bandwidth?

The A16 has 200 GB/s of memory bandwidth. Higher bandwidth helps with faster data transfer between GPU memory and compute cores.

What data types does the A16 support?

The A16 supports 6 precision formats. Training: BF16, FP16, TF32, FP32. Inference: INT4, INT8.

Does the A16 support NVLink?

No. The A16 is a PCIe-only GPU with no NVLink, so it is better suited to single-GPU inference and smaller-scale workloads than large distributed training jobs.

How much does the A16 cost per hour?

A16 pricing currently ranges from $0.47/hr to $0.56/hr per GPU, depending on the provider, instance type, and billing model.

How much does the A16 cost per month?

At 720 hours per month, one A16 can cost between $339.12 to $405.90 per month, depending on the provider. Reserved and spot pricing can lower that further.

Which cloud providers offer the A16?

The A16 is available from 3 cloud providers: Vultr, Sesterce, Runcrate.

Can I rent the A16 in the cloud?

Yes. We currently track 6 A16 listings across 3 cloud providers:

Billing type Listings Avg $/GPU/hr
On-demand 6 $0.52/hr

Technical Specifications

GPU Architecture NVIDIA Ampere architecture
GPU Memory 4x 16 GB GDDR6
Memory Bandwidth 4x 200 GB/s
Error-Correcting Code (ECC) Yes
NVIDIA Ampere architecture-based CUDA Cores 4x 1280
NVIDIA Third-Generation Tensor Cores 4x 40
NVIDIA Second-Generation RT Cores 4x 10
FP32 | TF32 | TF32' (TFLOPS) 4x 4.5, 4x 9, 4x 18
FP16 | FP16' (TFLOPS) 4x 17.9, 4x 35.9
INT8 | INT8' (TOPS) 4x 35.9, 4x 71.8
System Interface PCIe Gen4 (x16)
Max Power Consumption 250W
Thermal Solution Passive
Form Factor Full height, full length (FHFL) Dual Slot
Power Connector 8-pin CPU
Encode/Decode Engines 4 NVENC, 8 NVDEC (includes AV1 decode)
Secure and Measured Boot with Hardware Root of Trust for GPU Yes (optional)
vGPU Software Support NVIDIA Virtual PC (vPC), NVIDIA Virtual Applications (vApps), NVIDIA RTX Virtual Workstation (vWS), NVIDIA AI Enterprise, NVIDIA Virtual Compute Server (vCS)
Graphics APIs DirectX 12.07, Shader Model 5.17, OpenGL 4.68, Vulkan 1.18
Compute APIs CUDA, DirectCompute, OpenCL™, OpenACC®
MIG Support No

Source: official Nvidia A16 datasheet.

Alternatives to Nvidia A16

Last updated