Nvidia RTX 4000
Entry-level Turing professional card for visualization and light compute.
- Arch
- Turing
- Memory
- 8 GB GDDR6
- Bandwidth
- 416 GB/s
- Released
- Q4 2018
Price history
Weekly median price per GPU per hour
At a glance Updated 12 hours ago
- Cheapest
- $0.56 /GPU/hr
- on-demand · not verified in stock
- Median price (current)
- --
- 90-day trend
- --
- Coverage
- 1 provider
- 1 price point
Pricing & availability
How this list works
The two views
By provider shows one card per company, placed where its best offer ranked. By configuration lists every offer.
Order
What you can rent comes first: in stock, then waitlist, then not reported, then out of stock. Priced offers rank ahead of quote-only, and on-demand ahead of other billing types.
Within a group, five factors set the order:
- Location: datacenter proximity, blended with provider HQ.
- Price: hourly price, per GPU and in total.
- Billing type: reserved ahead of spot, spot ahead of quote-only.
- Specs: more VRAM, vCPUs and RAM.
- Provider diversity: a provider's repeat rows rank slightly lower, so one company can't take all the top spots.
Search
We match the provider, its country, GPU model, form factor, billing type, availability and instance name. Matching is partial and case-insensitive. Everyday words work too, so "interruptible" finds spot and "sold out" finds out of stock.
Transparency and funding
- Ads and sponsors: Paid placements sit at the top of the list and are always labeled as sponsored content. Sponsorship never influences the organic ranking itself.
- Affiliates: Affiliate links are marked. We may earn a commission if you click them, but commissions never affect the order.
- Prices: Shown in USD, converted at daily reference rates where a provider publishes in another currency. A month means 720 hours. How we estimate costs.
No providers match .
Heads up: A provider's own page may quote a different figure, for example on tax, region or a promotion. A monthly-only plan shows a derived hourly rate. Verify before provisioning. More on how we price.
What can the RTX 4000 run?
One RTX 4000 has 8 GB of VRAM. Even at 4-bit, the weights of most modern AI models exceed the card's usable VRAM, and context adds KV cache on top. Below are some open-weight LLMs, with the estimated memory and GPUs each one needs.
| Model | Memory (INT4 / FP4) | RTX 4000s needed | Cost /hr | Cost /mo |
|---|---|---|---|---|
|
|
19 GB
|
3
|
–
|
–
|
|
|
24 GB
|
4
|
–
|
–
|
|
|
68 GB
|
10
|
–
|
–
|
|
|
178 GB
|
25
|
–
|
–
|
|
|
241 GB
|
34
|
–
|
–
|
|
|
306 GB
|
43
|
–
|
–
|
|
|
1,545 GB
|
215
|
–
|
–
|
Estimates based on the median on-demand rate. Memory is weights plus FP16/BF16 KV cache at 32K context per request. GPU counts assume 90% of advertised VRAM is usable. No guarantee of runtime support, usable performance, or that a matching quantized build exists. How we estimate costs.
Technical Specifications
Nvidia RTX 4000 · Per GPU
| Compute · dense | |
|---|---|
| FP16 / BF16 | 57 TFLOPS |
| FP32 | 7.1 TFLOPS |
| Precision support | FP16FP32INT4INT8 |
| Memory | |
|---|---|
| Capacity | 8 GB GDDR6 |
| Bandwidth | 416 GB/s |
| Bus width | 256-bit |
| Silicon | |
|---|---|
| Architecture | Turing |
| Process | TSMC 12nm FFN |
| Shader cores | 2,304 CUDA cores |
| Matrix cores | 288 Tensor cores |
| Compute units | 36 SMs |
| Fabric and host | |
|---|---|
| Host interface | PCIe 3.0 x16 |
| Power | |
|---|---|
| Board power | 160 W |
| Cooling | Active |
Source: official Nvidia RTX 4000 datasheet.
Frequently Asked Questions
How much does the RTX 4000 cost per hour?
As of September 14, 2026, we track 1 config from 1 provider. Prices are per GPU per hour.
| Billing type | Configs | Cheapest |
|---|---|---|
On-demand | 1 | $0.56 |
No median for on-demand: we only show this when at least 3 providers list the GPU on that billing type.
Who has the cheapest RTX 4000?
The lowest listed on-demand price for the RTX 4000 is $0.56 per GPU per hour from Paperspace, though we haven't verified current stock.
Where can I rent an RTX 4000?
GetDeploying currently tracks RTX 4000 configs from 1 provider. See the full price comparison above for every provider and config.
How many RTX 4000 GPUs do I need?
More than one: even at 4-bit, the weights of most modern AI models exceed one RTX 4000's usable VRAM, and context adds KV cache on top. Larger models run across multiple GPUs: the model table above shows the estimated memory and GPU count for popular open-weight LLMs.
Why choose the RTX 4000?
Single-slot 160W card with Turing Tensor and RT cores, which is why it appears in dense workstation and cloud desktop fleets. Suits CAD, GPU rendering and light inference where the card has to fit a thin chassis and a small power budget.
When is the RTX 4000 not a good fit?
8GB of VRAM caps model size, and Turing predates BF16 and FP8, so mixed-precision work runs slower than on later cards. For AI workloads, A4000 doubles the memory in the same single-slot form factor. The RTX 4000 stays a fit for legacy CUDA applications, VDI and cloud desktop rendering.
Alternatives to Nvidia RTX 4000
Last updated