Gemini 2.5 Flash vs GPT-5 Mini
Data reconciled August 15, 2026
Short answer
Gemini 2.5 Flash and GPT-5 Mini each win on different axes — neither is simply cheaper or better across the board.
GPT-5 Mini is 1.2× cheaper on input.
GPT-5 Mini is 1.3× cheaper on output.
Gemini 2.5 Flash holds a 3.9× larger context window.
GPT-5 Mini returns a 2× longer answer per call.
Side by side
| Parameter | Gemini 2.5 Flash | GPT-5 Mini |
|---|---|---|
| Provider | Google (Gemini) | OpenAI |
| Input, $ per 1M tokens | $0.30 | $0.25 |
| Output, $ per 1M tokens | $2.50 | $2.00 |
| Cache read, $ per 1M | $0.03 | $0.025 |
| Context window | 1,048,576 | 272,000 |
| Max output | 65,535 | 128,000 |
| Vision | yes | yes |
| Tools | yes | yes |
| Reasoning | yes | yes |
| Status | available | available |
What it costs on your own workload
A price gap per million tokens says little until your volumes are plugged in: on a long fixed system prompt the winner is the model with cheap cache reads, not the one with a cheap input rate. Run Gemini 2.5 Flash and GPT-5 Mini against your own numbers.
Open the cost calculatorWhat is deliberately absent
We do not compare answer quality and we do not reprint third-party benchmarks. We have no quality measurements of our own, and external tables go stale faster than prices while being compiled, almost always, by someone with a stake in the result. What is here is only what we reconcile ourselves every day: prices, limits and supported capabilities.