Nvidia RTX 4000 SFF Ada

Nvidia RTX 4000 SFF Ada

Low-power professional card for small-form-factor workstations and dense inference nodes.

Latest update just now
Compare vs other GPUs →
Aggregating historical prices...

Key Specifications

Architecture
Ada Lovelace
Memory
20GB GDDR6
Memory Bandwidth
280 GB/s
Release date
Q2 2023

RTX 4000 SFF Ada Pricing and Availability

With a spread of only 12%, price optimization yields diminishing returns for the RTX 4000 SFF Ada. The market is relatively flat, with the lowest rate sitting at $0.44/hr per GPU (on-demand).

3 providers

How this list works

The two views

By provider shows one card per company, placed where its best offer ranked. By configuration lists every offer.

Order

What you can rent comes first: in stock, then waitlist, then not reported, then out of stock. Priced offers rank ahead of quote-only, and on-demand ahead of other billing types.

Within a group, five factors set the order:

  1. Location: datacenter proximity, blended with provider HQ.
  2. Price: hourly price, per GPU and in total.
  3. Billing type: reserved ahead of spot, spot ahead of quote-only.
  4. Specs: more VRAM, vCPUs and RAM.
  5. Provider diversity: a provider's repeat rows rank slightly lower, so one company can't take all the top spots.

Search

We match the provider, its country, GPU model, form factor, billing type, availability and instance name. Matching is partial and case-insensitive. Everyday words work too, so "interruptible" finds spot and "sold out" finds out of stock.

Transparency and funding

  • Ads and sponsors: Paid placements sit at the top of the list and are always labeled as sponsored content. Sponsorship never influences the organic ranking itself.
  • Affiliates: Affiliate links are marked. We may earn a commission if you click them, but commissions never affect the order.
  • Prices: Shown in USD, converted at daily reference rates where a provider publishes in another currency. A month means 720 hours. How we estimate costs.
Runpod logo

Runpod Our sponsor

USA 1 config 1x

On-Demand from $0.18
From $0.18 / GPU / hr On-Demand Visit website
Koyeb logo

Koyeb In stock

France 1 config 1x

On-Demand from $0.50
From $0.50 / GPU / hr On-Demand Visit website
Hetzner logo

Hetzner

Germany 1 config 1x

On-Demand from $0.44
From $0.44 / GPU / hr On-Demand Visit website

Heads up: We do our best to keep these specs & prices accurate. However, cloud costs may fluctuate based on region, usage, and other factors not listed here. These are estimates based on common setups and are for informational purposes only. Always verify current rates & exact specs with the provider before provisioning.

Frequently Asked Questions

Why choose the RTX 4000 SFF Ada?

20GB GDDR6 with ECC inside a 70W board power envelope, so it runs in small-form-factor workstations and dense chassis that cannot supply or cool a full-height card. Ada Lovelace FP8 Tensor Cores cover small-model inference, AI development, rendering and viewport work.

When is the RTX 4000 SFF Ada not a good fit?

The 70W envelope costs throughput: single-precision performance is roughly 30% below the full-height RTX 4000 Ada, which shares the same 20GB and pin count. There is no NVLink to pool memory across cards. For continuous cloud inference at a similar power draw, the L4 offers 24GB and higher memory bandwidth.

What size AI models can the RTX 4000 SFF Ada run?

With 20GB of VRAM, the RTX 4000 SFF Ada can typically run 7B to 13B models in FP16, or larger models in 4-bit quantized form.

How much VRAM does the RTX 4000 SFF Ada have?

The RTX 4000 SFF Ada has 20GB of VRAM. Multi-GPU setups increase total memory, but that memory is not automatically pooled across GPUs.

What is the RTX 4000 SFF Ada's memory bandwidth?

The RTX 4000 SFF Ada has 280 GB/s of memory bandwidth. Higher bandwidth helps with faster data transfer between GPU memory and compute cores.

What data types does the RTX 4000 SFF Ada support?

The RTX 4000 SFF Ada supports 7 precision formats. Training: BF16, FP16, TF32, FP32. Inference: FP8, INT4, INT8.

Does the RTX 4000 SFF Ada support NVLink?

No. The RTX 4000 SFF Ada is a PCIe-only GPU with no NVLink, so it is better suited to single-GPU inference and smaller-scale workloads than large distributed training jobs.

How much does the RTX 4000 SFF Ada cost per hour?

RTX 4000 SFF Ada pricing currently ranges from $0.44/hr to $0.50/hr per GPU, depending on the provider, instance type, and billing model.

How much does the RTX 4000 SFF Ada cost per month?

At 720 hours per month, one RTX 4000 SFF Ada can cost between $316.80 to $360.00 per month, depending on the provider. Reserved and spot pricing can lower that further.

Which cloud providers offer the RTX 4000 SFF Ada?

The RTX 4000 SFF Ada is available from 3 cloud providers: Hetzner, Koyeb, Runpod.

Can I rent the RTX 4000 SFF Ada in the cloud?

Yes. We currently track 3 RTX 4000 SFF Ada listings across 3 cloud providers:

Billing type Listings Avg $/GPU/hr
On-demand 3 $0.37/hr

Technical Specifications

GPU Memory 20GB GDDR6
Memory Interface 160-bit
Memory Bandwidth 280 GB/s
Error Correcting Code Yes
Architecture NVIDIA Ada Lovelace
CUDA Cores 6,144
Tensor Cores (Fourth-generation) 192
RT Cores (Third-generation) 48
Single-Precision Performance 19.2 TFLOPS
RT Core Performance 44.3 TFLOPS
Tensor Performance 306.8 TFLOPS
System Interface PCIe 4.0 x16
Power Consumption Total board power: 70W
Thermal Solution Active
Form Factor 2.7" H x 6.6" L, dual slot
Display Connectors 4x mini DisplayPort 1.4a
Encode/Decode Engines 2x encode, 2x decode (+AV1 encode and decode)
VR Ready Yes
Graphics APIs DirectX 12 Ultimate, Shader Model 6.7, OpenGL 4.6, Vulkan 1.3
Compute APIs CUDA 12.0, OpenCL 3.0, DirectCompute
NVIDIA NVLink® No

Source: official Nvidia RTX 4000 SFF Ada datasheet.

Alternatives to Nvidia RTX 4000 SFF Ada

Last updated