Google Gemini 3.1 Flash Lite is a Google multimodal model for high-volume, low-latency tasks, structured outputs, thinking, and model routing.
Capabilities, design details, and architectural traits
Gemini 3.1 Flash Lite is built for high-frequency tasks that need low latency and low cost. It is framed as a lightweight model for scale, with controls for thinking and structured task handling.
| Distinctive trait | What is documented |
|---|---|
| High-volume focus | Optimized for high-frequency, lightweight tasks and high-volume traffic. |
| Multimodal inputs | Supports text, image, video, audio, and PDF inputs. |
| Thinking control | Supports thinking, with selectable thinking levels in AI Studio and Vertex AI. |
| Task routing use | Documented for model routing, where it can classify task complexity. |
The model is described as a cost-effective option for straightforward work at scale. Documented uses include translation, transcription, data extraction, document processing, and routing tasks.
Its API documentation also lists function calling, structured outputs, code execution, file search, search grounding, URL context, and Google Maps grounding. The model is presented as a practical choice when speed, API cost, and throughput matter more than broad general-purpose depth.
Independent evaluations · Artificial Analysis
Evaluate specifications, pricing, and independent benchmark indices
| Model Details | |||
|---|---|---|---|
| General Info | |||
| Provider | Anthropic | Anthropic | |
| Release Date | March 3, 2026 | July 24, 2026 | June 9, 2026 |
| Knowledge Cutoff | - | May 2026 | - |
| Context & Limits | |||
| Context Window | 1.0M Best Context Window | 1M | 1M |
| Pricing (per 1M tokens) | |||
| Input Pricing | $0.25 Best Input Pricing | $5 | $10 |
| Output Pricing | $1.50 Best Output Pricing | $25 | $50 |
| Modalities | |||
| Inputs | textimagevideofileaudio | textimage | textimagefile |
| Outputs | text | text | text |
| Benchmarks (0-100) | |||
| Intelligence Index | 25.6 | 63.1 Best Intelligence Index | 62.1 |
| Coding Index | 34.7 | 78.0 Best Coding Index | 76.5 |
| Agentic Index | 6.5 | 59.2 Best Agentic Index | 56.6 |
Graduate-level reasoning and expert Q&A evaluation.
Extremely difficult logical reasoning and knowledge.
Logical reasoning over long context windows.
Independent evaluation data provided by Artificial Analysis. To view the latest benchmarks and full details, visit their official site.