Google
Released August 13, 2026

Gemini 3.7 Flash

Gemini 3.7 Flash by Google: production-ready model for complex coding, agentic workflows, multi-step execution with tunable thinking and 1M token context.

Inputs
Text
Image
File
Audio
Video
Outputs
Text

Model Overview

Capabilities, design details, and architectural traits

Gemini 3.7 Flash - Workhorse model for coding and agents

Gemini 3.7 Flash is a Flash-tier model from Google, positioned specifically for complex coding, agentic workflows, and reliable multi-step execution. It ships as a generally available, production-ready model with an emphasis on token efficiency and multimodal reasoning.

TraitDetail
Tunable thinking levelsSupports low, medium, and high thinking levels, with medium as the default
Context and output1M token context window and 64k max output tokens
Core positioningBuilt for complex coding, agentic workflows, and reliable multi-step execution
Updated Antigravity agentShips with an updated Antigravity agent integrated into the model release
Efficiency focusEmphasis on token efficiency and reliable code generation

Reasoning control

The tunable thinking levels let developers choose between low, medium, and high reasoning depth per request, with medium as the default. This gives explicit control over the trade-off between response thoroughness and token cost, directly relevant to the model's focus on coding and agent workflows where multi-step reasoning quality matters.

Benchmark Performance

Independent evaluations · Artificial Analysis

39.1%
Intelligence
76.1%
Coding Index
35.2%
Agentic Index

Accuracy & Capability Details

GPQA - Graduate Science94.5%
Humanity's Last Exam47.9%
SciCode - Scientific Coding57.2%
Long Context Reasoning81.7%

Compare Models Side-by-Side

Evaluate specifications, pricing, and independent benchmark indices

Model Details
General Info
ProviderGoogleAnthropicAnthropic
Release DateAugust 13, 2026September 22, 2026September 28, 2026
Knowledge Cutoff--Jun 2026
Context & Limits
Context Window
1.0M
Best Context Window
1M1M
Pricing (per 1M tokens)
Input Pricing
$0.75
Best Input Pricing
$4$2
Output Pricing
$3.75
Best Output Pricing
$20$10
Modalities
Inputs
textimagefileaudiovideo
textimagefile
textimagefile
Outputs
text
text
text
Benchmarks (0-100)
Intelligence Index39.1
57.6
Best Intelligence Index
56.0
Coding Index76.1--
Agentic Index35.2--
Gemini 3.7 Flash
Claude Opus 5.5
Claude Sonnet 5.5

Humanity's Last Exam

Extremely difficult logical reasoning and knowledge.

48%
Gemini 3.7 Flash
Humanity's Last Exam
Score: 48%
Gemini 3.7 Flash
61%
Claude Opus 5.5
Humanity's Last Exam
Score: 61%
Claude Opus 5.5
55%
Claude Sonnet 5.5
Humanity's Last Exam
Score: 55%
Claude Sonnet 5.5

Long Context Reasoning

Logical reasoning over long context windows.

82%
Gemini 3.7 Flash
Long Context Reasoning
Score: 82%
Gemini 3.7 Flash
85%
Claude Opus 5.5
Long Context Reasoning
Score: 85%
Claude Opus 5.5
83%
Claude Sonnet 5.5
Long Context Reasoning
Score: 83%
Claude Sonnet 5.5

SciCode Benchmark

Scientific coding and mathematical modeling.

57%
Gemini 3.7 Flash
SciCode Benchmark
Score: 57%
Gemini 3.7 Flash
67%
Claude Opus 5.5
SciCode Benchmark
Score: 67%
Claude Opus 5.5
61%
Claude Sonnet 5.5
SciCode Benchmark
Score: 61%
Claude Sonnet 5.5

Independent evaluation data provided by Artificial Analysis. To view the latest benchmarks and full details, visit their official site.

Explore more from Google

Other models by Google

Top AI Models

Leading alternatives by intelligence score

View all