Google Cloud In stock
USA 36 configs 1x-4x
Legacy inference card for lightweight, cost-sensitive deployments.
You can start using a P4 for $0.99/hr per GPU (on-demand). It's a low-risk price point for smaller workloads compared to the market high of $1.90/hr. Spot instances start lower, at $0.22/hr per GPU.
USA 36 configs 1x-4x
Heads up: We do our best to keep these specs & prices accurate. However, cloud costs may fluctuate based on region, usage, and other factors not listed here. These are estimates based on common setups and are for informational purposes only. Always verify current rates & exact specs with the provider before provisioning.
8GB GDDR5 at very low cost. INT8 inference support. Suitable for lightweight, cost-sensitive inference deployments.
Only 8GB VRAM on Pascal architecture with no Tensor Cores. Only viable for very lightweight legacy inference or basic video transcoding.
The median on-demand price across providers has risen about 220% since August 2025, from $0.36 to $1.17/hr per GPU. This reflects the market-wide median, which moves both when providers change prices and when lower or higher priced offerings enter the market.
With 8GB of VRAM, the P4 is mainly limited to small quantized models and lightweight GPU workloads.
The P4 has 8GB of VRAM. Multi-GPU setups increase total memory, but that memory is not automatically pooled across GPUs.
The P4 has 192 GB/s of memory bandwidth. Higher bandwidth helps with faster data transfer between GPU memory and compute cores.
The P4 supports 3 precision formats. Training: FP16, FP32. Inference: INT8.
No. The P4 is a PCIe-only GPU with no NVLink, so it is better suited to single-GPU inference and smaller-scale workloads than large distributed training jobs.
P4 pricing currently ranges from $0.22/hr to $1.90/hr per GPU, depending on the provider, instance type, and billing model.
At 720 hours per month, one P4 can cost between $160.96 to $1,370.30 per month, depending on the provider. Reserved and spot pricing can lower that further.
The P4 is available from 1 cloud provider: Google Cloud.
Yes. We currently track 36 P4 listings across 1 cloud providers:
| Billing type | Listings | Avg $/GPU/hr |
|---|---|---|
| On-demand | 9 | $1.26/hr |
| Reserved | 18 | $0.60/hr |
| Spot | 9 | $0.25/hr |
| Feature | Specification |
|---|---|
| GPU Architecture | NVIDIA Pascal™ |
| Single-Precision Performance | 5.5 TeraFLOPS* |
| Integer Operations (INT8) | 22 TOPS* (Tera-Operations per Second) |
| GPU Memory | 8 GB GDDR5 |
| Memory Bandwidth | 192 GB/s |
| System Interface | Low-Profile PCI Express Form Factor |
| Max Power | 75W |
| Enhanced Programmability | Yes |
| ECC Protection | Yes |
| Server-Optimized for Data Center | Yes |
| Hardware-Accelerated Video Engine | 1x Decode Engine, 2x Encode Engine |
| * With Boost Clock Enabled |
Source: official Nvidia P4 datasheet.
Last updated