Ask
Price change

fal raises H100 GPU list price to $4.50/h, matching H200

Fal pricing

fal's Serverless & Compute table now lists the H100 80GB GPU at a $4.50/h List Price, up from $3.99/h three weeks earlier — a 12.8% rise that erases the price gap to the H200. The negotiated 'as low as' floor held at $1.89/h.

Before

H100 80GB: List Price $3.99/h, As low as $1.89/h (set 2026-07-23 when fal split GPU pricing into List Price / As low as columns).

After

H100 80GB: List Price $4.50/h, As low as $1.89/h — now identical to H200's $4.50/h list price. All other GPU rates (H200, B200, B300, RTX PRO 6000) and every Model API rate are unchanged.

fal’s self-serve GPU rental got more expensive for one SKU. Less than three weeks after introducing a two-column “List Price” / “As low as” compute table — which itself doubled the effective self-serve price off the old single “$1.89/hr” headline — fal raised the H100’s List Price again, from $3.99/h to $4.50/h, erasing the gap that used to separate the 80GB H100 from the larger 141GB H200. The negotiated “As low as” floor for custom deployments held at $1.89/h, so the increase falls entirely on buyers who never contact sales for a lower rate. Every other rate on the page — H200, B200, B300, RTX PRO 6000, and the full Model APIs table for video and image — was unchanged between the July 22 and August 11 captures.

From Fal's pricing timeline
GPU compute split into List Price and 'As low as'

The Serverless & Compute table dropped its per-second column and now publishes two hourly rates per GPU. The rates fal used to advertise as the price became the floor (H100 $1.89/h, H200 $2.10/h) while new list prices sit roughly 2x higher (H100 $3.99/h, H200 $4.50/h), so a self-serve buyer who does not contact sales sees a materially higher number. B200 lost its 'contact us' gate at $6.25/h list ($3.49/h as low as); B300 288GB ($8.50/h, $4.49/h) and RTX PRO 6000 96GB ($2.99/h, $1.10/h) joined the fleet; A100 40GB was retired. Model API rates were unchanged.

About Fal
fal.ai ↗

fal (fal.ai) is a generative-media inference platform that prices purely on usage: serverless per-output model APIs plus dedicated GPU compute, with no seats, no subscriptions, and no free tier on its public pricing page.

Pricing model pure usage
Sales motion self servesales led
Free tier
No
Commits
None
Transparency
public

Fal pricing history

  1. Jul 2026
    GPU compute split into List Price and 'As low as'
  2. Jun 2026
    Serverless & Compute + Model APIs
  3. Jul 2025
    Modern layout: Output-Based Pricing + B200 'contact us'
  4. May 2025
    H100 cut to $1.89/hr; full per-hour fleet
  5. Apr 2025
    Per-hour GPU pricing introduced ($1.99/hr H100)
Full Fal timeline

More Fal activity

All pricing activity