← All comparisons

GPT-4o mini vs Gemini 2.5 Flash: token pricing & cost compared (July 2026)

How GPT-4o mini (OpenAI) and Gemini 2.5 Flash (Google) actually compare on price — input, output, and total monthly cost at the MAU you care about. Updated against live OpenRouter pricing.

OpenAI

GPT-4o mini

Input
$0.15 / 1M tokens
Output
$0.60 / 1M tokens
Output / Input ratio
4.0×
Context window
128K

Google

Gemini 2.5 Flash

Input
$0.30 / 1M tokens
Output
$2.50 / 1M tokens
Output / Input ratio
8.3×
Context window
1,048.576K

The short answer

GPT-4o mini is cheaper across the board for typical workloads.

Token pricing side-by-side

Input ($/1M)Output ($/1M)$0$0.65$1.3$1.95$2.6
  • GPT-4o mini
  • Gemini 2.5 Flash

Cost at your usage

GPT-4o mini

Per call

$0.00037

Per user / mo

$0.0375

Total / mo

$37

Gemini 2.5 Flash

Per call

$0.00140

Per user / mo

$0.1400

Total / mo

$140

GPT-4o mini is cheaper by $0.1025/user/month — a saving of $103/month at 1,000 MAU.

Total monthly bill across MAU

1005001K5K10K50K100K$0$4K$8K$12K$16K
  • GPT-4o mini
  • Gemini 2.5 Flash

Same usage profile applied to both models; only the per-token price differs.

Run this against your real numbers

The mini-calculator above is a preview. Use the full tools to model power-user blast radius, gross margin, and 12-month spend forecasts.

Comparing GPT-4o mini vs Gemini 2.5 Flash on price alone never tells the whole story — output-token weighting, context length, and your usage profile change the answer dramatically. Use the calculators below for the full picture.

Related tools: All comparisons · Cost Per User Calculator · How to calculate LLM cost per user