Google: Gemini 3.1 Flash Lite (batch)
Gemini 3.1 Flash Lite by Google is an AI model designed for handling extensive text inputs up to 1048576 tokens and supports multiple input modalities including text, images, videos, files, and audio. It offers robust reasoning capabilities and can engage with tools, making it versatile for various applications. While the model excels in flexibility and functionality, its benchmark coverage remains unproven, lacking independent validation scores. Given its pricing at $0.125 per million input tokens and $0.75 per million output tokens, Gemini 3.1 Flash Lite might be a cost-effective choice for organizations that need a broad-spectrum AI solution but should carefully consider the absence of benchmark data before deployment.
- Model ID
- google/gemini-3.1-flash-lite:batch
- Vendor
- Released
- May 2026
- Tokenizer
- Gemini
- Input Modalities
- text, image, video, file, audio
- Output Modalities
- text
- Max Output
- 65,536 tokens
- Tool Calling
- ✓ supported
- Structured Output
- ✓ supported
- Reasoning Mode
- ✓ supported
- Vision
- ✓ accepts images
- Audio
- ✓ accepts audio
- Moderated
- no
What it costs in practice
Computed from the current $0.12/M input and $0.75/M output rates. Run your own numbers →
| Job | Tokens | Cost |
|---|---|---|
| Summarize a 50-page report | 30k in / 1.5k out | under $0.01 |
| Classify 1,000 customer emails | 500k in / 50k out | $0.10 |
| A month of a busy support chatbot | 5M in / 2M out | $2.12 |
Category rankings
Where Google: Gemini 3.1 Flash Lite (batch) places across the 9 categories it ranks in. How we rank →
| # | Category | Score |
|---|---|---|
| #6 | TranscriptionVoice · of 25 ranked | 123 |
| #10 | TTS ReplacementVoice · of 25 ranked | 115 |
| #10 | Video Auto-TaggingVideo · of 25 ranked | 123 |
| #20 | Cheap Bulk InferenceCost · of 25 ranked | 137 |
| #20 | Real-Time ChatLatency · of 25 ranked | 117 |
| #21 | Social Media PostsWriting · of 25 ranked | 119 |
| #21 | Voice Assistant BackendVoice · of 25 ranked | 123 |
| #21 | Audio SummarizationVoice · of 25 ranked | 139 |
| #22 | Self-Hosted / LocalCost · of 25 ranked | 117 |
Similar models
Google: Gemini 3.6 Flash
Google: Gemini 3.6 Flash (batch)
Google: Gemini 3.5 Flash Lite
Google: Gemini 3.5 Flash Lite (batch)
Google: Gemini 3.5 Flash
Google: Gemini 3.5 Flash (batch)
Quick answers
- How much does Google: Gemini 3.1 Flash Lite (batch) cost?
- $0.12 per million input tokens and $0.75 per million output tokens.
- What is Google: Gemini 3.1 Flash Lite (batch)'s context window?
- 1,048,576 tokens, roughly 1,572 pages of text in a single request.
- Does Google: Gemini 3.1 Flash Lite (batch) support tool calling?
- It supports tool calling, structured output, a reasoning mode.
- Can Google: Gemini 3.1 Flash Lite (batch) process images?
- Yes, it accepts image input.