Skip to content

Gemini 2.5 Flash vs Gemini 2.5 Flash Lite

Data reconciled August 15, 2026

Short answer

Gemini 2.5 Flash Lite is 3× cheaper than Gemini 2.5 Flash on input and 6.3× cheaper on output, while giving up nothing on context or capabilities. On measurable parameters the choice is unambiguous.

Side by side

ParameterGemini 2.5 FlashGemini 2.5 Flash Lite
ProviderGoogle (Gemini)Google (Gemini)
Input, $ per 1M tokens$0.30$0.10
Output, $ per 1M tokens$2.50$0.40
Cache read, $ per 1M$0.03$0.01
Context window1,048,5761,048,576
Max output65,53565,535
Visionyesyes
Toolsyesyes
Reasoningyesyes
Statusavailableavailable

What it costs on your own workload

A price gap per million tokens says little until your volumes are plugged in: on a long fixed system prompt the winner is the model with cheap cache reads, not the one with a cheap input rate. Run Gemini 2.5 Flash and Gemini 2.5 Flash Lite against your own numbers.

Open the cost calculator

What is deliberately absent

We do not compare answer quality and we do not reprint third-party benchmarks. We have no quality measurements of our own, and external tables go stale faster than prices while being compiled, almost always, by someone with a stake in the result. What is here is only what we reconcile ourselves every day: prices, limits and supported capabilities.

Other comparisons

Model catalog