Fal.ai logo

Fal.ai

Serverless platform for running AI models

Founded in 2021, Fal.ai is a cloud platform for deploying AI models with a focus on inference for generative content. It allows developers to run and fine-tune models without managing complex infrastructure.

Example customers include PlayAI, Quora Poe, Genspark, Hedra.

Fal.ai Homepage

Fal.ai Homepage

What's good about Fal.ai

  • Optimized for fast inference, especially for generative media
  • Cost-effective, pay-as-you-go pricing model
  • Offers both serverless GPU instances

Fal.ai pricing examples

Fal.ai uses a usage-based pricing model, ensuring you only pay for the compute you consume. It offers two main structures:

  • GPU Pricing: Billed per second for deploying custom applications on their GPU fleet.
  • Output-Based Pricing: For models hosted by Fal.ai, billing is based on the output generated, such as per image, per megapixel, or per second of video.

Fal.ai GPUs

Here are some of the GPU configurations offered by Fal.ai:

5 GPU models

How this list works

Order

  • By GPU model: one card per GPU, the most looked-up models first, then the most widely offered ones.
  • By configuration: every configuration we track. What you can rent now comes first, and within that one row per GPU model before any repeats, so the opening rows cover the range rather than one model's sizes.

The price on a card

Each card leads with a per GPU per hour rate. We prefer configurations the provider reports as in stock, on-demand rates before reserved or spot, then take the lowest per-GPU rate among those.

Search

We match the GPU model, form factor, billing type, availability and instance name. Matching is partial and case-insensitive. Everyday words work too, so "interruptible" finds spot and "sold out" finds out of stock.

Transparency and funding

  • Ads and sponsors: Paid placements sit at the top of the list and are always labeled as sponsored content. Sponsorship never influences the organic ranking itself.
  • Affiliates: Affiliate links are marked. We may earn a commission if you click them, but commissions never affect the order.
  • Prices: Shown in USD, converted at daily reference rates where a provider publishes in another currency. A month means 720 hours. How we estimate costs.
Nvidia

Nvidia H100

80GB 2 configs 1x

On-Demand from $4.50 Custom from $1.89
From $4.50 / GPU / hr On-Demand Visit website
Nvidia

Nvidia B300

288GB 2 configs 1x

On-Demand from $8.50 Custom from $4.49
From $8.50 / GPU / hr On-Demand Visit website
Nvidia

Nvidia B200

192GB 2 configs 1x

On-Demand from $6.25 Custom from $3.49
From $6.25 / GPU / hr On-Demand Visit website
Nvidia

Nvidia H200

141GB 2 configs 1x

On-Demand from $4.50 Custom from $2.10
From $4.50 / GPU / hr On-Demand Visit website
Nvidia

Nvidia RTX PRO 6000

96GB 2 configs 1x

On-Demand from $2.99 Custom from $1.10
From $2.99 / GPU / hr On-Demand Visit website

Heads up: We do our best to keep these specs & prices accurate. However, cloud costs may fluctuate based on region, usage, and other factors not listed here. These are estimates based on common setups and are for informational purposes only. Always verify current rates & exact specs with the provider before provisioning. You can find Fal.ai's latest pricing here.

Which services does Fal.ai offer

Here are some of the services that Fal.ai offers:

Alternatives to Fal.ai

Compare Fal.ai against other cloud providers: