Google: Gemini 2.5 Flash Lite
Gemini 2.5 Flash Lite from Google is a robust AI model designed for handling extensive text input with a context length of 1MB and supports various modalities including text, images, files, audio, and video. It excels in reasoning tasks and can engage with tools to enhance its functionality, making it versatile for complex problem-solving. This model should be considered by users requiring comprehensive capabilities across multiple data types and strong reasoning abilities. Its price point of $0.1 per million input tokens and $0.4 per million output tokens makes it cost-effective, especially given its benchmark score of 9.9 across independent evaluations.
Benchmark results
Independent, published benchmarks. Blended score 9.9 across 1 benchmark, last refreshed 2026-07-23. How scoring works →
| Benchmark | Measures | Score |
|---|---|---|
| AI Index | broad capability composite | 11.4 |
- Model ID
- google/gemini-2.5-flash-lite
- Vendor
- Released
- July 2025
- Tokenizer
- Gemini
- Input Modalities
- text, image, file, audio, video
- Output Modalities
- text
- Max Output
- 65,535 tokens
- Tool Calling
- ✓ supported
- Structured Output
- ✓ supported
- Reasoning Mode
- ✓ supported
- Vision
- ✓ accepts images
- Audio
- ✓ accepts audio
- Moderated
- no
What it costs in practice
Computed from the current $0.10/M input and $0.40/M output rates. Run your own numbers →
| Job | Tokens | Cost |
|---|---|---|
| Summarize a 50-page report | 30k in / 1.5k out | under $0.01 |
| Classify 1,000 customer emails | 500k in / 50k out | $0.07 |
| A month of a busy support chatbot | 5M in / 2M out | $1.30 |
Category rankings
Where Google: Gemini 2.5 Flash Lite places across the 10 categories it ranks in. How we rank →
| # | Category | Score |
|---|---|---|
| #7 | TranscriptionVoice · of 22 ranked | 123 |
| #7 | Audio SummarizationVoice · of 22 ranked | 140 |
| #13 | Social Media PostsWriting · of 25 ranked | 119 |
| #13 | Voice Assistant BackendVoice · of 25 ranked | 123 |
| #13 | Cheap Bulk InferenceCost · of 25 ranked | 137 |
| #13 | Real-Time ChatLatency · of 25 ranked | 118 |
| #14 | Self-Hosted / LocalCost · of 25 ranked | 117 |
| #15 | TTS ReplacementVoice · of 22 ranked | 115 |
| #18 | Video SummarizationVideo · of 25 ranked | 140 |
| #20 | Code CompletionCode · of 25 ranked | 132 |
Similar models
Google: Gemini 3.6 Flash
Google: Gemini 3.5 Flash Lite
Google: Gemini 3.5 Flash
Google: Gemini 3.1 Flash Lite
Google: Gemma 4 26B A4B
Google: Gemma 4 26B A4B (free)
Quick answers
- How much does Google: Gemini 2.5 Flash Lite cost?
- $0.10 per million input tokens and $0.40 per million output tokens.
- What is Google: Gemini 2.5 Flash Lite's context window?
- 1,048,576 tokens, roughly 1,572 pages of text in a single request.
- Does Google: Gemini 2.5 Flash Lite support tool calling?
- It supports tool calling, structured output, a reasoning mode.
- Can Google: Gemini 2.5 Flash Lite process images?
- Yes, it accepts image input.