Ask
Launch

Gemini 3.8 Flash launches on Gemini 3.6/3.7's shared discounted rate

Google pricing

Google launched Gemini 3.8 Flash on Gemini 3.6/3.7's shared $0.75/$3.75 per-1M-token rate, plus new Transcribe models and Lyria 3.5.

Before

Gemini 3.6 and 3.7 Flash were the newest Flash-tier models, both on shared introductory pricing of $0.75/$3.75 per 1M input/output tokens through Dec 31, 2026 (reverting to $1.50/$7.50 Jan 1, 2027).

After

Gemini 3.8 Flash launches as the newest flagship Flash model, joining 3.6/3.7 Flash on the identical $0.75/$3.75 rate; Gemini 3.5 Transcribe and Transcribe Live (speech-to-text) and Lyria 3.5 (music) also launch, and Gemini 2.0 Flash, 2.0 Flash-Lite, Gemini 2.5 Flash-Lite Preview, Imagen 4, Veo 3, and Veo 2 are fully removed from the pricing page.

Google added Gemini 3.8 Flash to the Gemini API and Vertex AI, describing it on its own pricing-page listing as “Our most intelligent Flash model, engineered for long-horizon software engineering, autonomous agents, and complex enterprise workflows.” Rather than shipping at a new price point, Gemini 3.8 Flash joined Gemini 3.6 Flash and Gemini 3.7 Flash on the exact same discounted rate: Vertex AI’s on-page banner now reads “Gemini 3.8 Flash, Gemini 3.7 Flash, Gemini 3.6 Flash, and CodeMender using these models are offered with introductory pricing of $0.75 / $3.75 per 1M tokens input / output through December 31, 2026. Starting January 1, 2027, standard pricing of $1.5 / $7.5 per 1M tokens input / output will apply.” Non-global (regional) endpoints carry the existing +10% premium on both the introductory and standard rate.

The same update cycle added two dedicated speech-to-text models — Gemini 3.5 Transcribe ($2.00/1M input, $12.00/1M output) and Gemini 3.5 Transcribe Live ($3.50/1M input, $21.00/1M output) — promoted Gemini Omni Flash from Preview to general availability at an unchanged price, and added Lyria 3.5 ($0.08 per full song) as a new flagship music-generation model. In the same cycle, Google fully removed Gemini 2.0 Flash, Gemini 2.0 Flash-Lite, Gemini 2.5 Flash-Lite Preview, Imagen 4, Veo 3, and Veo 2 from the live pricing page — all past their previously-announced 2026 shutdown dates.

From Google's pricing timeline
Gemini 3.8 Flash launches; Transcribe models added; legacy SKUs retired

Google launched Gemini 3.8 Flash, its new "most intelligent Flash model" for long-horizon software engineering and autonomous agents, directly onto the same shared introductory pricing as Gemini 3.6 and 3.7 Flash: $0.75/1M input and $3.75/1M output (cached $0.075) through December 31, 2026, reverting to $1.50/$7.50 (cached $0.15) on January 1, 2027 — non-global endpoints carry the existing +10% premium. The same update added two dedicated speech-to-text models, Gemini 3.5 Transcribe ($2.00/$12.00 per 1M) and Gemini 3.5 Transcribe Live ($3.50/$21.00 per 1M); promoted Gemini Omni Flash from Preview to GA at an unchanged price; and added Lyria 3.5 ($0.08/song), with Lyria 3 relabeled a legacy family. Gemini 2.0 Flash, 2.0 Flash-Lite, Gemini 2.5 Flash-Lite Preview, Imagen 4, Veo 3, and Veo 2 were fully removed from the live pricing page, past their previously-announced shutdown dates. The consumer Google AI plans page added a bundled "Google Health Premium" perk (value undisclosed) for Pro/Ultra and a free-year-of-Plus student promotion; consumer USD list prices are unchanged, with $4.99/$19.99/$199.99 re-confirmed against an independent App Store listing and the $99.99 Ultra 5x tier carried forward from its last direct USD observation rather than freshly re-verified this cycle.

About Google
ai.google.dev ↗

Google Gemini API operates on a pure pay-per-token model with no subscription or seat fee — developers pay only for input and output tokens consumed, billed through Google Cloud.

Pricing model pure usagefreemium
Billing units tokensrequestsapi calls
Sales motion self serveplgsales led
Free tier
Yes
Commits
None
Transparency
public

Google pricing history

  1. Sep 2026
    Gemini 3.8 Flash launches; Transcribe models added; legacy SKUs retired
  2. Aug 2026
    Gemini 3.7 Flash launches; 3.6 Flash gets temporary 50% price cut
  3. Jul 2026
    Gemini Robotics ER 2 launches; Gemma 4 fine-tuning SKUs added
  4. Jul 2026
    Gemini 3.6 Flash & 3.5 Flash-Lite launched; Google AI Plus USD confirmed
  5. Jul 2026
    AlphaEvolve agent added to Vertex AI pricing
Full Google timeline

More Google activity

All pricing activity