Nvidia H100
In stock- On-Demand
- from $7.79
GPU cloud with per-second billing and an inference API
Runcrate homepage
Per-second billing on GPU compute with no minimum commitment: stop the instance to stop the meter. Inference is billed per token against a public rate card. Reserved and dedicated capacity is available with volume discounts on monthly commitments.
Each card leads with a per GPU per hour rate. We prefer configurations the provider reports as in stock, on-demand rates before reserved or spot, then take the lowest per-GPU rate among those.
We match the GPU model, form factor, billing type, availability and instance name. Matching is partial and case-insensitive. Everyday words work too, so "interruptible" finds spot and "sold out" finds out of stock.
Runcrate lists 14 GPU models across 16 configurations, including the H100, H200 and A100. Prices start at $0.20 /GPU/hr on demand.
No GPU models match .
Heads up: A provider's own page may quote a different figure, for example on tax, region or a promotion. A monthly-only plan shows a derived hourly rate. Verify before provisioning. More on how we price. See Runcrate's pricing.
Runcrate is a GPU cloud founded in 2024. It provides on-demand GPU instances with per-second billing, full root access, and an OpenAI-compatible inference API serving open-source models.
The platform aggregates GPU capacity across eight global regions spanning North America, Europe, and Asia. It targets AI teams that need to train, fine-tune, and serve models, offering both self-serve compute and reserved dedicated capacity.
Runcrate offers 1 of the 27 services we track.
| Service | Provider page |
|---|---|
| GPU-powered Servers | On Runcrate |
Compare Runcrate against other cloud providers.
Our data for Runcrate was last updated on Oct. 2, 2026.