Gemini 3.6 Flash gets a temporary 50% price cut as Gemini 3.7 Flash launches
Google launched Gemini 3.7 Flash and moved it and Gemini 3.6 Flash onto shared introductory pricing of $0.75/$3.75 per 1M tokens through Dec 31, 2026 — half of 3.6 Flash's prior $1.50/$7.50 rate, reverting Jan 1, 2027.
Gemini 3.6 Flash was priced flat at $1.50/1M input and $7.50/1M output tokens (cached input $0.15/1M), with no newer Flash-tier model above it.
Gemini 3.7 Flash launches as the new flagship Flash model; both 3.7 Flash and 3.6 Flash are billed at $0.75/1M input and $3.75/1M output (cached input $0.075/1M) through December 31, 2026, before standard $1.50/$7.50 pricing resumes January 1, 2027.
Google shipped Gemini 3.7 Flash, described as its “most capable Flash model for agentic workflows and multimodal reasoning,” and simultaneously placed the existing Gemini 3.6 Flash on the exact same discounted rate. An on-page banner on the Vertex AI pricing page states plainly: “Gemini 3.7 Flash and Gemini 3.6 Flash are offered with introductory pricing of $0.75/$3.75 per 1M tokens input/output through December 31, 2026. Starting January 1, 2027, standard pricing of $1.5/$7.5 per 1M tokens input/output will apply.” For Gemini 3.6 Flash specifically — a model that has been generally available since July 2026 at a flat $1.50/$7.50 rate — this amounts to a genuine, if temporary, 50% price cut rather than just an introductory rate for a brand-new model.
The same capture cycle turned up several other new Gemini API/Vertex AI surfaces: a Gemini Embedding 2 multimodal embedding model ($0.20/1M text plus per-modality image/audio/video rates); an explicit standalone rate for the Gemini Deep Research agent ($2.00/1M input, $12.00/1M output) where none previously existed; a new CodeMender vulnerability-patching agent billed at the exact underlying model rate with no surcharge (unlike AlphaEvolve’s 2× markup); and the replacement of the four Gemma 4 fine-tuning SKUs published in July 2026 with a single Gemma 4 26B Vertex serving price. Vertex AI’s Model Garden also grew a large “Partner models on Agent Platform” catalog spanning Anthropic, xAI, DeepSeek, MiniMax, Moonshot, Qwen, GLM, OpenAI, Meta, and Mistral AI.
Google launched Gemini 3.7 Flash, a new flagship Flash model, and simultaneously moved Gemini 3.6 Flash onto shared introductory pricing of $0.75/1M input and $3.75/1M output (cached input $0.075/1M) — through December 31, 2026, per an on-page Vertex AI banner. Standard pricing of $1.50/$7.50 (cached $0.15) resumes for both models on January 1, 2027, meaning 3.6 Flash's real price is temporarily halved from its prior flat $1.50/$7.50 rate. Same cycle: Gemini Embedding 2 (new multimodal embedding model) launched at $0.20/1M text plus per-modality image/audio/video rates; the Gemini Deep Research agent gained an explicit standalone rate ($2.00/1M input, $12.00/1M output) for the first time; a new CodeMender agent SKU appeared billed at the exact underlying model rate with no surcharge (unlike AlphaEvolve's 2×); the four Gemma 4 fine-tuning SKUs published 2026-07-30 were replaced by a Gemma 4 26B Vertex serving price ($0.15/$0.60 per 1M); and Vertex AI's Model Garden expanded its partner-model catalog (Anthropic, xAI, DeepSeek, MiniMax, Moonshot, Qwen, GLM, OpenAI, Meta Llama, Mistral AI).