Nvidia T4

Nvidia T4

Low-cost inference GPU with wide cloud availability.

Last update 7 minutes ago
Compare vs other GPUs →
Aggregating historical prices...

Key Specifications

Architecture
Turing
Memory per GPU
16 GB GDDR6
Memory bandwidth
300 GB/s
GPU interconnect
PCIe only
Release date
Q3 2018

T4 Pricing and Availability

Of 122 listings across 10 cloud providers, 94 are verified in stock. Cheapest of those: on-demand from $0.13 per GPU per hour, spot from $0.06, reservations from $0.21 (36 mo). The median on-demand price is $0.81, down 6% over the past 90 days.

10 providers

How this list works

The two views

By provider shows one card per company, placed where its best offer ranked. By configuration lists every offer.

Order

What you can rent comes first: in stock, then waitlist, then not reported, then out of stock. Priced offers rank ahead of quote-only, and on-demand ahead of other billing types.

Within a group, five factors set the order:

  1. Location: datacenter proximity, blended with provider HQ.
  2. Price: hourly price, per GPU and in total.
  3. Billing type: reserved ahead of spot, spot ahead of quote-only.
  4. Specs: more VRAM, vCPUs and RAM.
  5. Provider diversity: a provider's repeat rows rank slightly lower, so one company can't take all the top spots.

Search

We match the provider, its country, GPU model, form factor, billing type, availability and instance name. Matching is partial and case-insensitive. Everyday words work too, so "interruptible" finds spot and "sold out" finds out of stock.

Transparency and funding

  • Ads and sponsors: Paid placements sit at the top of the list and are always labeled as sponsored content. Sponsorship never influences the organic ranking itself.
  • Affiliates: Affiliate links are marked. We may earn a commission if you click them, but commissions never affect the order.
  • Prices: Shown in USD, converted at daily reference rates where a provider publishes in another currency. A month means 720 hours. How we estimate costs.
Vast.ai logo

Vast.ai Our sponsor

USA 5 configs 1x-2x

On-Demand from $0.13 Spot from $0.06
From $0.13 / GPU / hr On-Demand Visit website
Theta EdgeCloud logo

Theta EdgeCloud In stock

USA 2 configs 1x-2x

On-Demand from $0.20
From $0.20 / GPU / hr On-Demand Visit website
Zenlayer logo

Zenlayer In stock

USA 2 configs 1x-2x

On-Demand from $0.56
From $0.56 / GPU / hr On-Demand PCIe Visit website
Microsoft Azure logo

Azure In stock

USA 16 configs 1x-4x

On-Demand from $0.53 Reserved from $0.25 Spot from $0.10
From $0.53 / GPU / hr On-Demand Visit website
Amazon Web Services logo

AWS In stock

USA 21 configs 1x-8x

On-Demand from $0.53 Reserved from $0.27
From $0.53 / GPU / hr On-Demand Visit website
Google Cloud logo

Google Cloud In stock

USA 48 configs 1x-4x

On-Demand from $0.60 Reserved from $0.21 Spot from $0.22
From $0.60 / GPU / hr On-Demand Visit website
Cerebrium logo

Cerebrium

USA 4 configs 1x-8x

On-Demand from $0.59
From $0.59 / GPU / hr On-Demand Visit website
Replicate logo

Replicate

USA 1 config 1x

On-Demand from $0.81
From $0.81 / GPU / hr On-Demand Visit website
Alibaba Cloud logo

Alibaba Cloud

Singapore 12 configs 1x-4x

On-Demand from $1.12 Reserved from $0.62
From $1.12 / GPU / hr On-Demand Visit website
Leaseweb logo

Leaseweb

Netherlands 11 configs 1x-2x

Reserved from $0.23
From $0.23 / GPU / hr Reserved (1mo) Visit website

Heads up: Prices are estimates from published rates for common setups and vary by region and usage; verify with the provider before provisioning.

Frequently Asked Questions

Why choose the T4?

16GB GDDR6 in a 70W low-power form factor. Very low cost per hour across cloud providers. INT8 Tensor Cores for efficient inference on smaller models.

When is the T4 not a good fit?

Turing architecture lacks BF16 and FP8. 16GB VRAM limits model sizes. Slow for training. For production inference, the L4 offers significantly better throughput.

Are T4 prices going up or down?

The median on-demand price across providers has fallen about 6% over the past 90 days, from about $0.86 to $0.81 per GPU per hour.

How much does the T4 cost per month?

At 720 hours per month, one T4 costs about $583 at the median on-demand price. Cheapest verified in stock: $97 per month on-demand, $154 reserved (36 mo), $40 spot.

Which cloud providers offer the T4?

The T4 is listed by 10 providers. Cheapest verified in stock: Vast.ai, Theta EdgeCloud, Amazon Web Services and Microsoft Azure.

What data types does the T4 support?

The T4 supports 4 precision formats. Training: FP16, FP32. Inference: INT4, INT8.

Does the T4 support NVLink?

No. Multi-GPU setups communicate over PCIe.

What size AI models can the T4 run?

One T4 has 16 GB GDDR6. The table shows the smallest T4 node that fits each model, assuming moderate context, and what that node costs at the cheapest in-stock on-demand rate for one T4, $0.13 per hour.

Model 4-bit 8-bit BF16
2x (16 GB), from $0.26/hr 4x (32 GB), from $0.52/hr 8x (65 GB), from $1.04/hr
2x (18 GB), from $0.26/hr 4x (37 GB), from $0.52/hr 8x (74 GB), from $1.04/hr
4x (42 GB), from $0.52/hr 8x (84 GB), from $1.04/hr Over 8x (168 GB)
8x (70 GB), from $1.04/hr Over 8x (140 GB) Over 8x (281 GB)
Over 8x (141 GB) Over 8x (282 GB) Over 8x (564 GB)
Over 8x (452 GB) Over 8x (904 GB) Over 8x (1,807 GB)

Estimated memory required is parameters × bytes per weight at each precision, plus 20% for KV cache and runtime overhead. Anything larger than 8x needs more than one node.

How much does the T4 cost per hour?

As of August 27, 2026, the median on-demand price is $0.81 per GPU per hour across 9 providers.

Billing type Listings Median / GPU / hr Cheapest in stock
On-demand 41 $0.81 $0.13 (Vast.ai)
Reserved 63 $0.51 $0.21 (36 mo, Google Cloud)
Spot 18 $0.27 $0.06 (Vast.ai)

Technical Specifications

GPU Architecture NVIDIA Turing
NVIDIA Turing Tensor Cores 320
NVIDIA CUDA® Cores 2,560
Single-Precision 8.1 TFLOPS
Mixed-Precision (FP16/FP32) 65 TFLOPS
INT8 130 TOPS
INT4 260 TOPS
GPU Memory 16 GB GDDR6
Memory Bandwidth 300 GB/sec
ECC Yes
Interconnect Bandwidth 32 GB/sec
System Interface x16 PCIe Gen3
Form Factor Low-Profile PCIe
Thermal Solution Passive
Compute APIs CUDA, NVIDIA TensorRT™, ONNX

Source: official Nvidia T4 datasheet.

Alternatives to Nvidia T4

Last updated