Home › Blog › Gemini Flash price hike hits three models

Switching Gemini Flash Versions Won't Dodge the January Price Hike

8 October 2026 · AI & LLMs · 4 min read

We wrote about Gemini 3.7 Flash's price doubling back in September — $0.75/$3.75 per million tokens through December 31, 2026, then $1.50/$7.50 from January 1, 2027. Since then, Google shipped 3.8 Flash, and it's reasonable to assume a newer model got a cleaner rate card. It didn't. I pulled the current Gemini API pricing docs directly and checked every Flash generation Google currently sells: 3.6 Flash, 3.7 Flash and 3.8 Flash all carry the exact same intro-rate clause, on the exact same date. If your plan was "move to the newest Flash before January and dodge the hike," that plan doesn't work — there's nowhere in the current Flash lineup left to move to.

Every current Flash model, side by side

МодельNow, through Dec 31, 2026From Jan 1, 2027Изменять
Gemini 3.8 Flash$0.75 / $3.75$1.50 / $7.502.00x
Gemini 3.7 Flash$0.75 / $3.75$1.50 / $7.502.00x
Gemini 3.6 Flash$0.75 / $3.75$1.50 / $7.502.00x
Gemini 3.5 Flash (no scheduled change)$1.50 / $9.00$1.50 / $9.00none
Gemini 3.5 Flash-Lite$0.30 / $2.50$0.30 / $2.50none
Близнецы 2.5 Флэш (previous gen)$0.30 / $2.50$0.30 / $2.50none
Gemini 2.5 Flash-Lite$0.10 / $0.40$0.10 / $0.40none

Three model names, three separate launch dates, one identical price line. That's not a coincidence of similar models landing on similar numbers — it's the same promotional structure applied to an entire model family: cheap intro rate to pull in adoption, doubling on the same January 1 date regardless of which specific Flash build you picked. Google's docs don't call out 3.6 or 3.8 by name the way the 3.7 announcement got attention — the clause just sits under the same rate card line for all three.

What this actually costs on a real workload

A chatbot app running 300M input tokens and 60M output tokens a month — a mid-size support bot, not an edge case. Priced on whichever current-gen Flash model it happens to be running:

Input (300M)Output (60M)Итого за месяц
Now (3.6 / 3.7 / 3.8 Flash, through Dec 31)$225.00$225.00$450.00
From Jan 1, 2027 (same model, any of the three)$450.00$450.00$900.00
Gemini 3.5 Flash-Lite, any time (no hike)$90.00$150.00$240.00

The $450 jump is identical no matter which of the three current Flash generations the bot is running — there's no "upgrade to dodge it" move available inside the Flash tier itself. The only real way out is a tier change, not a version bump: Flash-Lite already runs this exact workload for $240/month today, before and after January, because it was never part of the promotional pricing in the first place.

Что бы я на самом деле сделал

Don't spend engineering time migrating between 3.6, 3.7 and 3.8 Flash for cost reasons — they're the same bill, just with different model weights attached. If the January number matters to your budget, the decision that actually changes it is Flash vs. Flash-Lite (or staying on 2.5 Flash), decided on an eval set, not a model-version number. Run your real traffic through Flash-Lite before December 31 and measure whether answer quality holds — if it does, the $210/month difference in the example above is free money; if it doesn't, you now know to budget for the doubled rate with your eyes open instead of discovering it on the January invoice.

Новичок в расценках LLM на основе использования? Начните с бесплатные руководства по стоимости API.

Разверните его самостоятельно: DigitalOcean — бесплатный кредит в размере 200 долларов США ↗ · Хостингер VPS ↗

Pricing verified against Google's official Gemini API pricing docs (ai.google.dev/gemini-api/docs/pricing), checked 2026-10-08. Reference estimates using Google's published pay-as-you-go rates — Vertex AI pricing, enterprise agreements and future rate changes vary; confirm current pricing at the source before budgeting. Workload figures (token counts) are illustrative scenarios computed against real published per-token rates, not vendor-supplied numbers.