Runpod Our sponsor
USA 4 configs 1x-2x
Legacy data center GPU with FP64 support for scientific computing.
Barriers to entry are low here. 21 cloud providers list the V100 with prices starting at just $0.09/hr per GPU (on-demand), making it accessible for testing without a large commitment. Spot instances start lower, at $0.05/hr per GPU.
USA 4 configs 1x-2x
USA 13 configs 1x-4x
USA 2 configs 1x-2x
USA 2 configs 1x
USA 48 configs 1x-8x
USA 6 configs 1x-8x
USA 6 configs 8x
France 4 configs 1x-4x
UAE 2 configs 1x-2x
Hong Kong 4 configs 1x-8x
USA 1 config 1x
USA 3 configs 1x-4x
USA 2 configs 1x
USA 3 configs 1x-8x
France 3 configs 1x-4x
Heads up: We do our best to keep these specs & prices accurate. However, cloud costs may fluctuate based on region, usage, and other factors not listed here. These are estimates based on common setups and are for informational purposes only. Always verify current rates & exact specs with the provider before provisioning.
32GB or 16GB HBM2 with 300 GB/s NVLink. Legacy data center GPU with wide availability and low cost. FP64 support makes it useful for scientific computing.
Volta architecture lacks BF16, TF32, and FP8 support. Significantly slower than Ampere or Hopper for AI workloads. Only consider if budget is the primary constraint.
The median on-demand price has held steady around $0.97/hr per GPU across providers.
With 32GB of VRAM, the V100 can typically run 7B to 13B models in FP16, or larger models in 4-bit quantized form.
The V100 has 32GB of VRAM. Multi-GPU setups increase total memory, but that memory is not automatically pooled across GPUs.
The V100 has 900 GB/s of memory bandwidth. Higher bandwidth helps with faster data transfer between GPU memory and compute cores.
The V100 supports 4 precision formats. Training: FP16, FP32. Inference: INT8. Scientific: FP64.
Yes. The V100 supports NVLink with 300 GB/s of bidirectional bandwidth. This helps accelerate multi-GPU communication.
V100 pricing currently ranges from $0.05/hr to $4.22/hr per GPU, depending on the provider, instance type, and billing model.
At 720 hours per month, one V100 can cost between $39.20 to $3,040.78 per month, depending on the provider. Reserved and spot pricing can lower that further.
The V100 is available from 21 cloud providers, including Verda, Amazon Web Services, Exoscale. Pricing and availability vary by region and billing model.
Yes. We currently track 175 V100 listings across 21 cloud providers:
| Billing type | Listings | Avg $/GPU/hr |
|---|---|---|
| On-demand | 75 | $1.57/hr |
| Reserved | 77 | $1.12/hr |
| Spot | 23 | $0.23/hr |
| V100 PCIe | V100 SXM2 | V100S PCIe | |
|---|---|---|---|
| GPU Architecture | NVIDIA Volta | NVIDIA Volta | NVIDIA Volta |
| NVIDIA Tensor Cores | 640 | 640 | 640 |
| NVIDIA CUDA® Cores | 5,120 | 5,120 | 5,120 |
| Double-Precision Performance | 7 TFLOPS | 7.8 TFLOPS | 8.2 TFLOPS |
| Single-Precision Performance | 14 TFLOPS | 15.7 TFLOPS | 16.4 TFLOPS |
| Tensor Performance | 112 TFLOPS | 125 TFLOPS | 130 TFLOPS |
| GPU Memory | 32 GB / 16 GB HBM2 | 32 GB HBM2 | 32 GB HBM2 |
| Memory Bandwidth | 900 GB/sec | 900 GB/sec | 1134 GB/sec |
| ECC | Yes | Yes | Yes |
| Interconnect Bandwidth | 32 GB/sec | 300 GB/sec | 32 GB/sec |
| System Interface | PCIe Gen3 | NVIDIA NVLink™ | PCIe Gen3 |
| Form Factor | PCIe Full Height/Length | SXM2 | PCIe Full Height/Length |
| Max Power Consumption | 250 W | 300 W | 250 W |
| Thermal Solution | Passive | Passive | Passive |
| Compute APIs | CUDA, DirectCompute, OpenCL™, OpenACC® | CUDA, DirectCompute, OpenCL™, OpenACC® | CUDA, DirectCompute, OpenCL™, OpenACC® |
Source: official Nvidia V100 datasheet.
Ada Lovelace inference GPU. From $0.20/hr per GPU across 32 providers.
Hopper datacenter GPU. From $0.60/hr per GPU across 52 providers.
Consumer Ada Lovelace GPU. From $0.14/hr per GPU across 16 providers.
Consumer Ampere gaming GPU. From $0.06/hr per GPU across 8 providers.
Turing inference GPU. From $0.06/hr per GPU across 9 providers.
Last updated