Vast.ai In stock
USA 2 configs 1x
Budget legacy GPU when VRAM matters more than speed.
The P40 is available at 1 cloud provider for $0.11/hr per GPU. Spot instances start lower, at $0.07/hr per GPU.
USA 2 configs 1x
Heads up: We do our best to keep these specs & prices accurate. However, cloud costs may fluctuate based on region, usage, and other factors not listed here. These are estimates based on common setups and are for informational purposes only. Always verify current rates & exact specs with the provider before provisioning.
24GB GDDR5 at very low cost. Legacy option for workloads that need more VRAM than the P4. Some availability from budget cloud providers.
Pascal architecture with no mixed precision support. GDDR5 bandwidth is very low. Only viable for legacy batch inference where VRAM matters more than speed.
The median on-demand price across providers has fallen about 95% since August 2025, from $2.07 to $0.11/hr per GPU. This reflects the market-wide median, which moves both when providers change prices and when lower or higher priced offerings enter the market.
With 24GB of VRAM, the P40 can typically run 7B to 13B models in FP16, or larger models in 4-bit quantized form.
The P40 has 24GB of VRAM. Multi-GPU setups increase total memory, but that memory is not automatically pooled across GPUs.
The P40 has 346 GB/s of memory bandwidth. Higher bandwidth helps with faster data transfer between GPU memory and compute cores.
The P40 supports 3 precision formats. Training: FP16, FP32. Inference: INT8.
No. The P40 is a PCIe-only GPU with no NVLink, so it is better suited to single-GPU inference and smaller-scale workloads than large distributed training jobs.
P40 pricing currently ranges from $0.07/hr to $0.11/hr per GPU, depending on the provider, instance type, and billing model.
At 720 hours per month, one P40 can cost between $48.67 to $77.47 per month, depending on the provider. Reserved and spot pricing can lower that further.
The P40 is available from 1 cloud provider: Vast.ai.
Yes. We currently track 2 P40 listings across 1 cloud providers:
| Billing type | Listings | Avg $/GPU/hr |
|---|---|---|
| On-demand | 1 | $0.11/hr |
| Spot | 1 | $0.07/hr |
| GPU | 1 NVIDIA Pascal GPU |
| CUDA Cores | 3,840 |
| Memory Size | 24 GB GDDR5 |
| H.264 1080p30 streams | 24 |
| Max vGPU instances | 24 (1 GB Profile) |
| vGPU Profiles | 1 GB, 2 GB, 3 GB, 4 GB, 6 GB, 8 GB, 12 GB, 24 GB |
| Form Factor | PCIe 3.0 Dual Slot (rack servers) |
| Power | 250 W |
| Thermal | Passive |
Source: official Nvidia P40 datasheet.
Last updated