Google Gemini 3.1 Pro Preview is a Google model for better thinking, reliable multi-step tool use, structured outputs, and software engineering.
Capabilities, design details, and architectural traits
Gemini 3.1 Pro Preview is built to refine the Gemini 3 Pro series for harder tasks. It emphasizes better thinking, more grounded answers, and reliable multi-step execution.
| Distinctive trait | What is documented |
|---|---|
| Software engineering focus | Optimized for software engineering behavior and usability. |
| Reliable multi-step work | Built for agentic workflows that require precise tool use and reliable multi-step execution. |
| Grounded responses | Described as more grounded and factually consistent. |
| Thinking and tools | Supports thinking, code execution, file search, function calling, search grounding, structured outputs, URL context, and Google Maps grounding. |
The model is presented as a preview step for improving the Gemini 3 Pro series. Its documented pattern is to make advanced reasoning practical in real-world domains rather than only producing a direct answer.
The docs also list support for text, image, video, audio, and PDF inputs, plus Batch API, Flex inference, and Priority inference. A separate custom-tools endpoint is documented for bash and custom tool workflows.
Independent evaluations · Artificial Analysis
Evaluate specifications, pricing, and independent benchmark indices
| Model Details | |||
|---|---|---|---|
| General Info | |||
| Provider | Anthropic | Anthropic | |
| Release Date | February 19, 2026 | July 24, 2026 | June 9, 2026 |
| Knowledge Cutoff | - | May 2026 | - |
| Context & Limits | |||
| Context Window | 1.0M Best Context Window | 1M | 1M |
| Pricing (per 1M tokens) | |||
| Input Pricing | $2 Best Input Pricing | $5 | $10 |
| Output Pricing | $12 Best Output Pricing | $25 | $50 |
| Modalities | |||
| Inputs | audiofileimagetextvideo | textimage | textimagefile |
| Outputs | text | text | text |
| Benchmarks (0-100) | |||
| Intelligence Index | 47.7 | 63.1 Best Intelligence Index | 62.1 |
| Coding Index | 68.8 | 78.0 Best Coding Index | 76.5 |
| Agentic Index | 23.0 | 59.2 Best Agentic Index | 56.6 |
Graduate-level reasoning and expert Q&A evaluation.
Extremely difficult logical reasoning and knowledge.
Logical reasoning over long context windows.
Independent evaluation data provided by Artificial Analysis. To view the latest benchmarks and full details, visit their official site.