Paperspace
USA 1 config 1x
Entry-level Turing professional card for visualization and light compute.
The RTX 4000 is available at 1 cloud provider for $0.56/hr per GPU.
USA 1 config 1x
Heads up: We do our best to keep these specs & prices accurate. However, cloud costs may fluctuate based on region, usage, and other factors not listed here. These are estimates based on common setups and are for informational purposes only. Always verify current rates & exact specs with the provider before provisioning.
Single-slot 160W card with Turing Tensor and RT cores, which is why it appears in dense workstation and cloud desktop fleets. Suits CAD, GPU rendering and light inference where the card has to fit a thin chassis and a small power budget.
8GB of VRAM caps model size, and Turing predates BF16 and FP8, so mixed-precision work runs slower than on later cards. For AI workloads, A4000 doubles the memory in the same single-slot form factor. The RTX 4000 stays a fit for legacy CUDA applications, VDI and cloud desktop rendering.
With 8GB of VRAM, the RTX 4000 is mainly limited to small quantized models and lightweight GPU workloads.
The RTX 4000 has 8GB of VRAM. Multi-GPU setups increase total memory, but that memory is not automatically pooled across GPUs.
The RTX 4000 has 416 GB/s of memory bandwidth. Higher bandwidth helps with faster data transfer between GPU memory and compute cores.
The RTX 4000 supports 4 precision formats. Training: FP16, FP32. Inference: INT4, INT8.
No. The RTX 4000 is a PCIe-only GPU with no NVLink, so it is better suited to single-GPU inference and smaller-scale workloads than large distributed training jobs.
RTX 4000 pricing is currently $0.56/hr per GPU.
At 720 hours per month, one RTX 4000 currently costs about $403.20/mo.
The RTX 4000 is available from 1 cloud provider: Paperspace.
Yes. We currently track 1 RTX 4000 listings across 1 cloud providers:
| Billing type | Listings | Avg $/GPU/hr |
|---|---|---|
| On-demand | 1 | $0.56/hr |
| GPU Memory | 8 GB GDDR6 |
| Memory Interface | 256-bit |
| Memory Bandwidth | Up to 416 GB/s |
| NVIDIA CUDA Cores | 2,304 |
| NVIDIA Tensor Cores | 288 |
| NVIDIA RT Cores | 36 |
| Single-Precision Performance | 7.1 TFLOPS |
| Tensor Performance | 57.0 TFLOPS |
| NVIDIA NVLink | Not supported |
| System Interface | PCI Express 3.0 x16 |
| Power Consumption | Total board power: 160 W, Total graphics power: 125 W |
| Thermal Solution | Active |
| Form Factor | 4.4" H x 9.5" L, Single Slot |
| Display Connectors | 3x DP 1.4, 1x USB-C (VirtualLink) |
| Max Simultaneous Displays | 4x 3840 x 2160 @ 120 Hz, 4x 5120 x 2880 @ 60 Hz, 2x 7680 x 4320 @ 60 Hz |
| VR Ready | Yes |
| Graphics APIs | DirectX 12.0, Shader Model 5.1, OpenGL 4.6, Vulkan 1.1 |
| Compute APIs | CUDA, DirectCompute, OpenCL™ |
Source: official Nvidia RTX 4000 datasheet.
Single-slot Ada Lovelace professional card with 20GB. From $0.16/hr per GPU across 7 providers.
Single-slot Ampere professional card with 16GB. From $0.03/hr per GPU across 11 providers.
Dual-slot Turing professional card with 16GB and NVLink. From $0.60/hr per GPU across 2 providers.
Single-slot Pascal professional card with 8GB. From $0.06/hr per GPU across 2 providers.
Last updated