fal splits GPU pricing into list vs 'as low as' rates and adds B300
fal's Serverless & Compute table now shows two hourly rates per GPU — a List Price and an 'As low as' rate. H100 lists at $3.99/h against the $1.89/h headline. B300 and RTX PRO 6000 join; A100 is gone.
Single per-hour and per-second rate per GPU: A100 40GB $0.99/h, H100 80GB $1.89/h, H200 141GB $2.10/h, B200 184GB 'contact us'.
Two published hourly rates per GPU (List Price / As low as): RTX PRO 6000 96GB $2.99/h / $1.10/h, H100 80GB $3.99/h / $1.89/h, H200 141GB $4.50/h / $2.10/h, B200 180GB $6.25/h / $3.49/h, B300 288GB $8.50/h / $4.49/h.
Proof of change
Fal's pricing pages, as we captured them on two dates.
Showing the whole page as captured — scroll either panel, or open it at full size.
Show 1 other page we compared
Showing the whole page as captured — scroll either panel, or open it at full size.
- These captures are 51 days apart, so anything that changed and reverted in between would not appear here.
Both images are our own captures, taken on the dates shown. The pixels are unmodified — nothing is retouched, and where a region is highlighted the marker is drawn over the image, not into it. Open either capture to see it at full resolution.
fal rebuilt the compute half of its pricing page. The Serverless & Compute table previously carried one rate per GPU shown at two granularities (per hour and per second); it now carries one granularity (per hour) at two price points — a List Price and an “As low as” rate reserved for custom deployments, with the footer routing buyers to [email protected].
The practical effect is a disclosed discount spread rather than a straight increase. The rates fal used to advertise as the price are now the floor: H100 at $1.89/h and H200 at $2.10/h survive intact in the “As low as” column, while the new List Price column sits roughly 2x higher ($3.99/h and $4.50/h). A self-serve buyer who does not contact sales now sees a materially higher number than they did a month ago.
The fleet itself also moved. B200 is no longer sales-gated — it went from “contact us” to $6.25/h list ($3.49/h as low as), and its VRAM is now listed as 180GB rather than 184GB. B300 (288GB, $8.50/h list, $4.49/h as low as) and RTX PRO 6000 (96GB, $2.99/h list, $1.10/h as low as) are new, and the RTX PRO 6000 replaces the removed A100 40GB as the cheapest entry point. Model API pricing was unchanged in this move: Wan 2.5 $0.05/s, Kling 2.5 Turbo Pro $0.07/s, Veo 3 $0.4/s, Ovi $0.2/video, Seedream V4 $0.03/image, Flux Kontext Pro $0.04/image, Nanobanana $0.0398/image, Qwen $0.02/MP.
The Serverless & Compute table dropped its per-second column and now publishes two hourly rates per GPU. The rates fal used to advertise as the price became the floor (H100 $1.89/h, H200 $2.10/h) while new list prices sit roughly 2x higher (H100 $3.99/h, H200 $4.50/h), so a self-serve buyer who does not contact sales sees a materially higher number. B200 lost its 'contact us' gate at $6.25/h list ($3.49/h as low as); B300 288GB ($8.50/h, $4.49/h) and RTX PRO 6000 96GB ($2.99/h, $1.10/h) joined the fleet; A100 40GB was retired. Model API rates were unchanged.