Fireworks AI raises $1.5B Series D at a $17.5B valuation
Fireworks AI raised a $1.5B Series D at a $17.5B valuation to expand enterprise AI inference infrastructure.
Quartz and Ventureburn (2026-07-16) reported Fireworks AI’s $1.5B Series D at a $17.5B valuation, earmarked for enterprise inference infrastructure expansion.
Inference-serving is the most price-competitive segment in the corpus; capital at this scale typically precedes per-token rate pressure rather than increases. No rate change was confirmed on the pricing page in this run.
Serverless inference is now documented as three named serving paths — Standard (default), Priority (higher reliability under peak traffic, set via service_tier) and Fast (100+ tokens/sec, selected by model ID) — replacing the earlier "Turbo + Priority" framing, with per-model input / cached-input / output rates published for each. Fireworks also shipped Fire Pass, an experimental promo-code pass that removes per-token charges on included open-weight models for personal agentic coding, and the site banner now announces a Series D and $1B ARR. Headline rate card (H100/H200 $7.00/hr, B200 $10.00/hr, B300 $12.00/hr, fine-tuning from $0.50 per 1M training tokens, batch at 50%) is unchanged.