Gemini Flash Lite
Google (Gemini)
At a glance
- Price per 1M tokens (input)
- $0.10
- Price per 1M tokens (output)
- $0.40
- Context
- 1,048,576 tokens
- Free access
- yes (see below)
Refreshed daily; data verified August 15, 2026. Published August 12, 2026.
The Gemini Flash Lite model is designed for businesses and developers who need to integrate image understanding and step-by-step reasoning capabilities into their applications. This model is well-suited for tasks that require analyzing visual data, making decisions based on that analysis, and taking subsequent actions. The vendor's approach to model development focuses on creating flexible and adaptable models that can be fine-tuned for specific use cases. Tool calling, or the ability to invoke external functions, is a key feature of this model, allowing it to interact with other systems and services. By leveraging these capabilities, developers can build more sophisticated and automated workflows.
Specifications & pricing
| Input (per 1M tokens) | $0.10 |
|---|---|
| Output (per 1M tokens) | $0.40 |
| Cache read (per 1M tokens) | $0.01 |
| Context window | 1,048,576 tokens |
| Max output | 65,535 tokens |
| Capabilities | images, tool calling, reasoning |
LiteLLM community dataset (MIT), verified August 15, 2026. Official Google (Gemini) pricing.
What Gemini Flash Lite would cost on your workload — run it through the cost calculator →
Where to try Gemini Flash Lite for free
- Google (Gemini) offers a free chat — Gemini (free plan). A vendor's free chat may run a different model from the same family — the exact model is not guaranteed.
- 🎁 On our promo-codes page: Gemini for free: access and discounts.
Frequently asked questions
What is the Gemini Flash Lite model good for+
The Gemini Flash Lite model is good for applications that require image understanding, such as image classification, object detection, and image segmentation, as well as tasks that involve step-by-step reasoning and decision-making
How do I get started with the Gemini Flash Lite model+
To get started with the Gemini Flash Lite model, you will need to review the model's documentation and familiarize yourself with its capabilities and limitations, then design and implement a workflow that leverages the model's features, such as tool calling and image understanding
How does the Gemini Flash Lite model differ from other models in the Gemini line+
The Gemini Flash Lite model differs from other models in the Gemini line in terms of its specific capabilities and focus areas, with this model emphasizing image understanding and step-by-step reasoning, while other models may focus on different areas, such as natural language processing or text analysis
What are the limitations of the Gemini Flash Lite model+
The limitations of the Gemini Flash Lite model include its potential difficulty in handling certain types of images or scenarios, such as low-quality or distorted images, and its reliance on the quality and accuracy of the data used to train it, which can impact its overall performance and effectiveness
Compare with others

Org chart: how to move your company onto AI
A practical map: which company roles and processes AI agents can take over, where to start, and in what order to roll it out.