GPT-5 nano vs Gemini 3.1 Flash: price comparison
Real 2026 API prices side by side, with the total cost of a typical 10M-input / 3M-output monthly workload.
Price & cost at a glance
| GPT-5 nano | Gemini 3.1 Flash | |
|---|---|---|
| Input / 1M tokens | $0.1 | $0.1 |
| Output / 1M tokens | $0.4 | $0.4 |
| Context window | 400K | 1000K |
| Cost — 10M in + 3M out | $2.2 | $2.2 |
Green = cheaper / larger. Prices per 1M tokens, USD. Snapshot 2026 — verify with the provider.
Which is cheaper?
Run your own token mix in the cost calculator, or see each model in depth: GPT-5 nano · Gemini 3.1 Flash.
Frequently asked
Is GPT-5 nano or Gemini 3.1 Flash cheaper?
At a typical 10M input + 3M output tokens per month, GPT-5 nano costs about $2.2 versus $2.2 for Gemini 3.1 Flash — GPT-5 nano is roughly 1.0x cheaper on this workload. Your ratio of input to output tokens changes the gap.
What's the price difference between GPT-5 nano and Gemini 3.1 Flash?
GPT-5 nano is $0.1/$0.4 per 1M input/output tokens; Gemini 3.1 Flash is $0.1/$0.4. Output tokens usually dominate a bill, so compare the output price first.
Which has the bigger context window, GPT-5 nano or Gemini 3.1 Flash?
GPT-5 nano supports about 400K tokens and Gemini 3.1 Flash about 1000K. A bigger window costs more per call because every token in context is billed.
Educational estimates — not affiliated with any provider.