cheapest capable
cheapest / month
flagship spread / mo

Full price table — your workload

Reference $/1M tokens (July 2026), monthly cost on your numbers above.

ProviderModelIn $/1MOut $/1MCost / mo

Cheapest by use case

⚠️ Reference prices, July 2026 — all three providers change pricing regularly. Confirm on each provider's pricing page. Output is billed separately and costs more than input on every model. · Report outdated price →
✓ Last verified: 2026-08-15· Source: official provider pricing page· Auto-monitored — report change →

How to actually choose

Brand loyalty is the most expensive habit in AI. The three providers leapfrog each other constantly, and the real cost difference comes from two things you control: the model tier you pick and how long your outputs are. A frontier model on short answers can be cheaper than a "cheap" model on rambling ones. Price the workload, not the logo — the table above does exactly that on your own numbers.

Drill into one provider with the GPT-4o calculator, the Claude calculator or the Gemini calculator, compare all models at once on the full AI API cost calculator, or estimate a whole product with the AI app cost estimator.

Three levers cut the bill further once you've picked a model, and all three providers offer all three. Prompt caching reprices repeated system prompts and documents at a fraction of the normal input rate — worth it for any app that resends the same instructions or context on every call; size it on the prompt caching savings calculator. A batch API gives roughly 50% off for jobs that can wait a few hours instead of needing a live response; see the batch API savings calculator. And for high-volume apps, routing routine requests to a cheaper model while reserving the flagship for hard cases is often the single biggest lever of all — model it on the AI model router cost calculator.

One more difference that doesn't show up in the price table: only Gemini has a genuine free API tier (Google AI Studio, no card required), which makes it the cheapest way to prototype before committing spend. OpenAI and Anthropic both require billing set up from the first request, though new accounts get limited trial credit.

FAQ

Which is cheapest: OpenAI, Claude or Gemini? For the very cheapest capable model, Gemini 2.0 Flash ($0.10 / $0.40 per 1M tokens) and GPT-4o mini ($0.15 / $0.60) lead. Among flagships, Gemini 2.5 Pro has the lowest input price, GPT-4o sits in the middle, and Claude Opus is the most expensive — the winner depends on the tier you need, not the brand.

Is Claude more expensive than GPT-4o? Claude Sonnet 4 ($3 / $15) is a little pricier than GPT-4o ($2.50 / $10), and Claude Opus 4 is far more expensive — but Claude Haiku 3.5 undercuts GPT-4o. Compare on the workload you actually run, since output length dominates the bill.

Do all three support prompt caching and batch discounts? Yes — all three offer caching for repeated prompts and a batch API at roughly half price. The real savings depend on how repetitive your prompts are and whether your job can tolerate the batch delay.

Does any of the three have a free API tier? Gemini is the only one — Google AI Studio gives free access with rate limits generous enough for prototyping. OpenAI and Anthropic require billing set up first, though both give limited trial credit to new accounts.