Gemini 3.8 Flash launches on Gemini 3.6/3.7's shared discounted rate
Google launched Gemini 3.8 Flash on Gemini 3.6/3.7's shared $0.75/$3.75 per-1M-token rate, plus new Transcribe models and Lyria 3.5.
Gemini 3.6 and 3.7 Flash were the newest Flash-tier models, both on shared introductory pricing of $0.75/$3.75 per 1M input/output tokens through Dec 31, 2026 (reverting to $1.50/$7.50 Jan 1, 2027).
Gemini 3.8 Flash launches as the newest flagship Flash model, joining 3.6/3.7 Flash on the identical $0.75/$3.75 rate; Gemini 3.5 Transcribe and Transcribe Live (speech-to-text) and Lyria 3.5 (music) also launch, and Gemini 2.0 Flash, 2.0 Flash-Lite, Gemini 2.5 Flash-Lite Preview, Imagen 4, Veo 3, and Veo 2 are fully removed from the pricing page.
Google added Gemini 3.8 Flash to the Gemini API and Vertex AI, describing it on its own pricing-page listing as “Our most intelligent Flash model, engineered for long-horizon software engineering, autonomous agents, and complex enterprise workflows.” Rather than shipping at a new price point, Gemini 3.8 Flash joined Gemini 3.6 Flash and Gemini 3.7 Flash on the exact same discounted rate: Vertex AI’s on-page banner now reads “Gemini 3.8 Flash, Gemini 3.7 Flash, Gemini 3.6 Flash, and CodeMender using these models are offered with introductory pricing of $0.75 / $3.75 per 1M tokens input / output through December 31, 2026. Starting January 1, 2027, standard pricing of $1.5 / $7.5 per 1M tokens input / output will apply.” Non-global (regional) endpoints carry the existing +10% premium on both the introductory and standard rate.
The same update cycle added two dedicated speech-to-text models — Gemini 3.5 Transcribe ($2.00/1M input, $12.00/1M output) and Gemini 3.5 Transcribe Live ($3.50/1M input, $21.00/1M output) — promoted Gemini Omni Flash from Preview to general availability at an unchanged price, and added Lyria 3.5 ($0.08 per full song) as a new flagship music-generation model. In the same cycle, Google fully removed Gemini 2.0 Flash, Gemini 2.0 Flash-Lite, Gemini 2.5 Flash-Lite Preview, Imagen 4, Veo 3, and Veo 2 from the live pricing page — all past their previously-announced 2026 shutdown dates.
Google launched Gemini 3.8 Flash, its new "most intelligent Flash model" for long-horizon software engineering and autonomous agents, directly onto the same shared introductory pricing as Gemini 3.6 and 3.7 Flash: $0.75/1M input and $3.75/1M output (cached $0.075) through December 31, 2026, reverting to $1.50/$7.50 (cached $0.15) on January 1, 2027 — non-global endpoints carry the existing +10% premium. The same update added two dedicated speech-to-text models, Gemini 3.5 Transcribe ($2.00/$12.00 per 1M) and Gemini 3.5 Transcribe Live ($3.50/$21.00 per 1M); promoted Gemini Omni Flash from Preview to GA at an unchanged price; and added Lyria 3.5 ($0.08/song), with Lyria 3 relabeled a legacy family. Gemini 2.0 Flash, 2.0 Flash-Lite, Gemini 2.5 Flash-Lite Preview, Imagen 4, Veo 3, and Veo 2 were fully removed from the live pricing page, past their previously-announced shutdown dates. The consumer Google AI plans page added a bundled "Google Health Premium" perk (value undisclosed) for Pro/Ultra and a free-year-of-Plus student promotion; consumer USD list prices are unchanged, with $4.99/$19.99/$199.99 re-confirmed against an independent App Store listing and the $99.99 Ultra 5x tier carried forward from its last direct USD observation rather than freshly re-verified this cycle.