Google: Gemini 2.5 Flash Lite
The Google Gemini 2.5 Flash Lite is an AI model that supports multiple input modalities, including text, image, file, audio, and video inputs, allowing users to interact with it in various ways. It can handle long contexts of up to 1 megabyte in length and provides tools and reasoning capabilities, though it does not output structured data. For those looking for a model that offers a balance between input flexibility and pricing, the Google Gemini 2.5 Flash Lite is worth considering, despite its premium cost. With a blended benchmark score of 0.8 across one independent benchmark, its performance stands out at this price point; however, it remains to be seen how it will perform in additional evaluation contexts due to limited coverage. Those willing to invest $0.1 per megabyte for input and $0.4 per megabyte for output may find the Google Gemini 2.5 Flash Lite suitable for their needs.
Benchmark results
Independent, published benchmarks. Blended score 0.8 across 1 benchmark, last refreshed 2026-09-07. How scoring works →
| Benchmark | Measures | Score |
|---|---|---|
| AI Index | broad capability composite | 2.3 |
- Model ID
- google/gemini-2.5-flash-lite
- Vendor
- Released
- July 2025
- Tokenizer
- Gemini
- Input Modalities
- text, image, file, audio, video
- Output Modalities
- text
- Max Output
- 65,535 tokens
- Tool Calling
- ✓ supported
- Structured Output
- ✓ supported
- Reasoning Mode
- ✓ supported
- Vision
- ✓ accepts images
- Audio
- ✓ accepts audio
- Moderated
- no
What it costs in practice
Computed from the current $0.10/M input and $0.40/M output rates. Run your own numbers →
| Job | Tokens | Cost |
|---|---|---|
| Summarize a 50-page report | 30k in / 1.5k out | under $0.01 |
| Classify 1,000 customer emails | 500k in / 50k out | $0.07 |
| A month of a busy support chatbot | 5M in / 2M out | $1.30 |
Category rankings
Where Google: Gemini 2.5 Flash Lite places across the 3 categories it ranks in. How we rank →
| # | Category | Score |
|---|---|---|
| #22 | TranscriptionVoice · of 25 ranked | 123 |
| #22 | Real-Time ChatLatency · of 25 ranked | 118 |
| #24 | Cheap Bulk InferenceCost · of 25 ranked | 137 |
Similar models
Google: Gemini 3.8 Flash
Google: Gemini 3.8 Flash (batch)
Google: Gemini 3.7 Flash
Google: Gemini 3.7 Flash (batch)
Google: Gemini 3.6 Flash
Google: Gemini 3.6 Flash (batch)
Quick answers
- How much does Google: Gemini 2.5 Flash Lite cost?
- $0.10 per million input tokens and $0.40 per million output tokens.
- What is Google: Gemini 2.5 Flash Lite's context window?
- 1,048,576 tokens, roughly 1,572 pages of text in a single request.
- Does Google: Gemini 2.5 Flash Lite support tool calling?
- It supports tool calling, structured output, a reasoning mode.
- Can Google: Gemini 2.5 Flash Lite process images?
- Yes, it accepts image input.