Fal.ai logo

Fal.ai

Serverless inference for generative image, video and audio models

HQ
United States of America flagUSA
Founded
2021
SOC 2
Type II
Fal.ai homepage Fal.ai homepage

At a glance GPU prices checked 1 month ago

Cheapest GPU
$1.10 /GPU/hr
RTX PRO 6000 · custom
H100
$1.89 /GPU/hr
custom
GPU models
5
10 configurations

What's good about Fal.ai

  • Optimized for fast inference, especially for generative media
  • Cost-effective, pay-as-you-go pricing model
  • Offers both serverless GPU instances

Pricing & billing

Pricing page

Fal.ai uses a usage-based pricing model, ensuring you only pay for the compute you consume. It offers two main structures:

  • GPU Pricing: Billed per second for deploying custom applications on their GPU fleet.
  • Output-Based Pricing: For models hosted by Fal.ai, billing is based on the output generated, such as per image, per megapixel, or per second of video.

Fal.ai GPUs

How this list works

Order

  • By GPU model: one card per GPU, the most looked-up models first, then the most widely offered ones.
  • By configuration: every configuration we track. What you can rent now comes first, and within that one row per GPU model before any repeats, so the opening rows cover the range rather than one model's sizes.

The price on a card

Each card leads with a per GPU per hour rate. We prefer configurations the provider reports as in stock, on-demand rates before reserved or spot, then take the lowest per-GPU rate among those.

Search

We match the GPU model, form factor, billing type, availability and instance name. Matching is partial and case-insensitive. Everyday words work too, so "interruptible" finds spot and "sold out" finds out of stock.

Transparency and funding

  • Ads and sponsors: Paid placements sit at the top of the list and are always labeled as sponsored content. Sponsorship never influences the organic ranking itself.
  • Affiliates: Affiliate links are marked. We may earn a commission if you click them, but commissions never affect the order.
  • Prices: Shown in USD, converted at daily reference rates where a provider publishes in another currency. A month means 720 hours. Pricing methodology.

Fal.ai lists 5 GPU models across 10 configurations, including the H100, B300 and B200. Prices start at $1.10 /GPU/hr custom.

Nvidia

2 configs 1x

On-Demand
from $4.50
Custom
from $1.89
From $4.50 /GPU/hr On-Demand
Visit website
Nvidia

2 configs 1x

On-Demand
from $8.50
Custom
from $4.49
From $8.50 /GPU/hr On-Demand
Visit website
Nvidia

2 configs 1x

On-Demand
from $6.25
Custom
from $3.49
From $6.25 /GPU/hr On-Demand
Visit website
Nvidia

2 configs 1x

On-Demand
from $4.50
Custom
from $2.10
From $4.50 /GPU/hr On-Demand
Visit website
Nvidia

2 configs 1x

On-Demand
from $2.99
Custom
from $1.10
From $2.99 /GPU/hr On-Demand
Visit website

Heads up: A provider's own page may quote a different figure, for example on tax, region or a promotion. A monthly-only plan shows a derived hourly rate. Verify before provisioning. More on how we price. See Fal.ai's pricing.

About Fal.ai

Founded in 2021, Fal.ai is a cloud platform for deploying AI models with a focus on inference for generative content. It allows developers to run and fine-tune models without managing complex infrastructure.

Example customers include PlayAI, Quora Poe, Genspark, Hedra.

Services

Fal.ai offers 1 of the 27 services we track.

Service Provider page
GPU-powered Servers On Fal.ai

Compliance

Fal.ai covers 2 of the 6 compliance frameworks we track.

Framework Level Checked Source
SOC 2 Type II trust.fal.ai
GDPR fal.ai

Heads up: Each row links to the provider's own statement, checked on the date shown. An attestation can cover one tier, one region or the platform alone rather than every service. Verify the scope before you rely on it.

Alternatives to Fal.ai

Compare Fal.ai against other cloud providers.

Our data for Fal.ai was last updated on Sept. 18, 2026.