Gemini 3.7 Flash
Gemini 3.7 Flash by Google: production-ready model for complex coding, agentic workflows, and multi-step execution with tunable thinking and 1M token context.
Model Overview
Capabilities, design details, and architectural traits
Gemini 3.7 Flash - Most intelligent workhorse model yet for coding and agents
Gemini 3.7 Flash is Google's most capable Flash-tier model, positioned specifically for complex coding, agentic workflows, and reliable multi-step execution. It ships as a generally available, production-ready model with an emphasis on greater token efficiency and stronger multimodal reasoning compared to its predecessor.
| Trait | Detail |
|---|---|
| Tunable thinking levels | Supports low, medium, and high thinking levels, with medium as the default |
| Context and output | 1M token context window and 64k max output tokens |
| Core positioning | Built for complex coding, agentic workflows, and reliable multi-step execution |
| Updated Antigravity agent | Ships with an updated Antigravity agent integrated into the model release |
| Efficiency focus | Greater token efficiency and more reliable code generation than prior Flash models |
Reasoning control
The tunable thinking levels let developers choose between low, medium, and high reasoning depth per request, with medium as the default. This gives explicit control over the trade-off between response thoroughness and token cost, directly relevant to the model's focus on coding and agent workflows where multi-step reasoning quality matters.
Benchmark Performance
Independent evaluations · Artificial Analysis
Accuracy & Capability Details
Compare Models Side-by-Side
Evaluate specifications, pricing, and independent benchmark indices
| Model Details | |||
|---|---|---|---|
| General Info | |||
| Provider | Anthropic | OpenAI | |
| Release Date | August 13, 2026 | September 1, 2026 | September 3, 2026 |
| Knowledge Cutoff | - | - | Apr 2026 |
| Context & Limits | |||
| Context Window | - | - | 1.1M |
| Pricing (per 1M tokens) | |||
| Input Pricing | $0.75 Best Input Pricing | $10 | $10 |
| Output Pricing | $3.75 Best Output Pricing | $50 | $50 |
| Modalities | |||
| Inputs | text | textimagefile | textimage |
| Outputs | text | text | text |
| Benchmarks (0-100) | |||
| Intelligence Index | 45.2 | 56.8 Best Intelligence Index | 54.7 |
| Coding Index | 76.1 | 81.6 Best Coding Index | 76.9 |
| Agentic Index | 36.6 | 58.2 Best Agentic Index | 51.6 |
GPQA Benchmark
Graduate-level reasoning and expert Q&A evaluation.
Humanity's Last Exam
Extremely difficult logical reasoning and knowledge.
Long Context Reasoning
Logical reasoning over long context windows.
Independent evaluation data provided by Artificial Analysis. To view the latest benchmarks and full details, visit their official site.