Home › Blog › Gemini Flash price hike hits three models

Switching Gemini Flash Versions Won't Dodge the January Price Hike

8 October 2026 · AI & LLMs · 4 min read

We wrote about Gemini 3.7 Flash's price doubling back in September — $0.75/$3.75 per million tokens through December 31, 2026, then $1.50/$7.50 from January 1, 2027. Since then, Google shipped 3.8 Flash, and it's reasonable to assume a newer model got a cleaner rate card. It didn't. I pulled the current Gemini API pricing docs directly and checked every Flash generation Google currently sells: 3.6 Flash, 3.7 Flash and 3.8 Flash all carry the exact same intro-rate clause, on the exact same date. If your plan was "move to the newest Flash before January and dodge the hike," that plan doesn't work — there's nowhere in the current Flash lineup left to move to.

Every current Flash model, side by side

ModelNow, through Dec 31, 2026From Jan 1, 2027Change
Gemini 3.8 Flash$0.75 / $3.75$1.50 / $7.502.00x
Gemini 3.7 Flash$0.75 / $3.75$1.50 / $7.502.00x
Gemini 3.6 Flash$0.75 / $3.75$1.50 / $7.502.00x
Gemini 3.5 Flash (no scheduled change)$1.50 / $9.00$1.50 / $9.00none
Gemini 3.5 Flash-Lite$0.30 / $2.50$0.30 / $2.50none
Gemini 2.5 Flash (previous gen)$0.30 / $2.50$0.30 / $2.50none
Gemini 2.5 Flash-Lite$0.10 / $0.40$0.10 / $0.40none

Three model names, three separate launch dates, one identical price line. That's not a coincidence of similar models landing on similar numbers — it's the same promotional structure applied to an entire model family: cheap intro rate to pull in adoption, doubling on the same January 1 date regardless of which specific Flash build you picked. Google's docs don't call out 3.6 or 3.8 by name the way the 3.7 announcement got attention — the clause just sits under the same rate card line for all three.

What this actually costs on a real workload

A chatbot app running 300M input tokens and 60M output tokens a month — a mid-size support bot, not an edge case. Priced on whichever current-gen Flash model it happens to be running:

Input (300M)Output (60M)Monthly total
Now (3.6 / 3.7 / 3.8 Flash, through Dec 31)$225.00$225.00$450.00
From Jan 1, 2027 (same model, any of the three)$450.00$450.00$900.00
Gemini 3.5 Flash-Lite, any time (no hike)$90.00$150.00$240.00

The $450 jump is identical no matter which of the three current Flash generations the bot is running — there's no "upgrade to dodge it" move available inside the Flash tier itself. The only real way out is a tier change, not a version bump: Flash-Lite already runs this exact workload for $240/month today, before and after January, because it was never part of the promotional pricing in the first place.

What I'd actually do

Don't spend engineering time migrating between 3.6, 3.7 and 3.8 Flash for cost reasons — they're the same bill, just with different model weights attached. If the January number matters to your budget, the decision that actually changes it is Flash vs. Flash-Lite (or staying on 2.5 Flash), decided on an eval set, not a model-version number. Run your real traffic through Flash-Lite before December 31 and measure whether answer quality holds — if it does, the $210/month difference in the example above is free money; if it doesn't, you now know to budget for the doubled rate with your eyes open instead of discovering it on the January invoice.

New to usage-based LLM pricing? Start with the free API-cost guides.

Deploy it yourself: DigitalOcean — $200 free credit ↗ · Hostinger VPS ↗

Pricing verified against Google's official Gemini API pricing docs (ai.google.dev/gemini-api/docs/pricing), checked 2026-10-08. Reference estimates using Google's published pay-as-you-go rates — Vertex AI pricing, enterprise agreements and future rate changes vary; confirm current pricing at the source before budgeting. Workload figures (token counts) are illustrative scenarios computed against real published per-token rates, not vendor-supplied numbers.