Replicate logo

Replicate

Model API with a community library, billed per second

HQ
United States of America flagUSA
Founded
2019
Replicate homepage Replicate homepage

At a glance GPU prices checked 4 hours ago

Cheapest GPU
$0.81 /GPU/hr
T4 · on-demand
H100
$5.49 /GPU/hr
on-demand
GPU models
5
17 configurations

What's good about Replicate

  • Run AI/ML models without having to manage the infrastructure
  • Excellent web UI to explore and try out models without code
  • Train and deploy custom models using Cog

Pricing & billing

Pricing page

Replicate uses a metered billing model. You only pay as long as your code is running, billed by the second based on the hardware selected.

Replicate GPUs

How this list works

Order

  • By GPU model: one card per GPU, the most looked-up models first, then the most widely offered ones.
  • By configuration: every configuration we track. What you can rent now comes first, and within that one row per GPU model before any repeats, so the opening rows cover the range rather than one model's sizes.

The price on a card

Each card leads with a per GPU per hour rate. We prefer configurations the provider reports as in stock, on-demand rates before reserved or spot, then take the lowest per-GPU rate among those.

Search

We match the GPU model, form factor, billing type, availability and instance name. Matching is partial and case-insensitive. Everyday words work too, so "interruptible" finds spot and "sold out" finds out of stock.

Transparency and funding

  • Ads and sponsors: Paid placements sit at the top of the list and are always labeled as sponsored content. Sponsorship never influences the organic ranking itself.
  • Affiliates: Affiliate links are marked. We may earn a commission if you click them, but commissions never affect the order.
  • Prices: Shown in USD, converted at daily reference rates where a provider publishes in another currency. A month means 720 hours. Pricing methodology.

Replicate lists 5 GPU models across 17 configurations, including the H100, H200 and A100. Prices start at $0.81 /GPU/hr on demand.

Nvidia

4 configs 1x-8x

On-Demand
from $5.49
From $5.49 /GPU/hr On-Demand
Visit website
Nvidia

4 configs 1x-8x

On-Demand
from $5.49
From $5.49 /GPU/hr On-Demand
Visit website
Nvidia

4 configs 1x-8x

On-Demand
from $5.04
From $5.04 /GPU/hr On-Demand
Visit website
Nvidia

4 configs 1x-8x

On-Demand
from $3.51
From $3.51 /GPU/hr On-Demand
Visit website
Nvidia

1 config 1x

On-Demand
from $0.81
From $0.81 /GPU/hr On-Demand
Visit website

Heads up: A provider's own page may quote a different figure, for example on tax, region or a promotion. A monthly-only plan shows a derived hourly rate. Verify before provisioning. More on how we price. See Replicate's pricing.

About Replicate

Founded in 2019 and acquired by Cloudflare in 2025, Replicate is a platform to train and deploy ML models in the cloud. You can run and fine tune open source models such as Meta's Llama 3, Mistral and Stable Diffusion without having to set up complex infrastructure.

They also feature a collection of generative models shared by the community, which you can run via the web UI or using their REST API.

Example customers include BuzzFeed, Labelbox, PhotoAI, Character.ai, Magnific, Unsplash, HeadshotPro.

Services

Replicate offers 1 of the 27 services we track.

Service Provider page
GPU-powered Servers On Replicate

Alternatives to Replicate

Compare Replicate against other cloud providers.

Our data for Replicate was last updated on Sept. 30, 2026.