Gemini 3.5 Flash-Lite is the fastest 3.5-class model, optimized for high-volume agentic workflows, subagent tasks, and document parsing with 1M context.
Capabilities, design details, and architectural traits
Gemini 3.5 Flash-Lite is a low-latency, natively multimodal reasoning model built on Gemini 3.1 Flash-Lite and positioned as the speed and cost leader within the Gemini 3.5 class. It is optimized for high-volume agentic workflows, subagent tasks, and document parsing where latency and API cost are the primary constraints.
| Trait | Detail |
|---|---|
| Throughput | 350 output tokens per second (Artificial Analysis Index), making it the fastest 3.5-class model |
| Subagent and document focus | Optimized for high-throughput, low-cost execution of subagent tasks and document parsing |
| Agentic workflow gains | Significantly outperforms prior Flash-Lite generations in agentic workflows |
| Multimodal inputs | Text, image, video, audio, and PDF with a 1M token context window; 64K token text output |
| Agentic capabilities | Supports function calling, file search, computer use (preview), thinking, URL context, and structured outputs |
| Not supported | Audio generation, image generation, and Live API are not available |
| Foundation | Based on Gemini 3.1 Flash-Lite architecture |
The model is distributed across Gemini App, Google AI Studio, Gemini Enterprise Agent Platform, and the Gemini API. It supports caching, batch API, flex inference, and priority inference to manage cost and throughput at production scale.
Independent evaluations · Artificial Analysis
Evaluate specifications, pricing, and independent benchmark indices
| Model Details | |||
|---|---|---|---|
| General Info | |||
| Provider | Anthropic | Anthropic | |
| Release Date | July 21, 2026 | July 24, 2026 | June 9, 2026 |
| Knowledge Cutoff | - | May 2026 | - |
| Context & Limits | |||
| Context Window | 1.0M Best Context Window | 1M | 1M |
| Pricing (per 1M tokens) | |||
| Input Pricing | $0.30 Best Input Pricing | $5 | $10 |
| Output Pricing | $2.50 Best Output Pricing | $25 | $50 |
| Modalities | |||
| Inputs | textimagefileaudiovideo | textimage | textimagefile |
| Outputs | text | text | text |
| Benchmarks (0-100) | |||
| Intelligence Index | 37.4 | 63.1 Best Intelligence Index | 62.1 |
| Coding Index | 49.3 | 78.0 Best Coding Index | 76.5 |
| Agentic Index | 27.2 | 59.2 Best Agentic Index | 56.6 |
Graduate-level reasoning and expert Q&A evaluation.
Extremely difficult logical reasoning and knowledge.
Logical reasoning over long context windows.
Independent evaluation data provided by Artificial Analysis. To view the latest benchmarks and full details, visit their official site.