← All comparisons

GPT-4o mini vs Gemini 2.0 Flash: token pricing & cost compared (July 2026)

How GPT-4o mini (OpenAI) and Gemini 2.0 Flash (Google) actually compare on price — input, output, and total monthly cost at the MAU you care about. Updated against live OpenRouter pricing.

OpenAI

GPT-4o mini

Input
$0.15 / 1M tokens
Output
$0.60 / 1M tokens
Output / Input ratio
4.0×
Context window
128K

Google

Gemini 2.0 Flash

Input
$0.00 / 1M tokens
Output
$0.00 / 1M tokens
Output / Input ratio
Context window

The short answer

Gemini 2.0 Flash is cheaper across the board for typical workloads.

Token pricing side-by-side

Input ($/1M)Output ($/1M)$0$0.15$0.3$0.45$0.6
  • GPT-4o mini
  • Gemini 2.0 Flash

Cost at your usage

GPT-4o mini

Per call

$0.00037

Per user / mo

$0.0375

Total / mo

$37

Gemini 2.0 Flash

Per call

$0.00000

Per user / mo

$0.0000

Total / mo

$0

Gemini 2.0 Flash is cheaper by $0.0375/user/month — a saving of $37/month at 1,000 MAU.

Total monthly bill across MAU

1005001K5K10K50K100K$0$950$2K$3K$4K
  • GPT-4o mini
  • Gemini 2.0 Flash

Same usage profile applied to both models; only the per-token price differs.

Run this against your real numbers

The mini-calculator above is a preview. Use the full tools to model power-user blast radius, gross margin, and 12-month spend forecasts.

Comparing GPT-4o mini vs Gemini 2.0 Flash on price alone never tells the whole story — output-token weighting, context length, and your usage profile change the answer dramatically. Use the calculators below for the full picture.

Related tools: All comparisons · Cost Per User Calculator · How to calculate LLM cost per user