DeepSeek V3.1 by DeepSeek is a 671B/37B-active hybrid MoE model with togglable thinking and non-thinking modes, optimized tool-use, and extended 128K context. MIT licensed.
Capabilities, design details, and architectural traits
DeepSeek V3.1 is the first model in the DeepSeek V3 family to unify thinking and non-thinking behavior in a single model. It is post-trained on top of DeepSeek V3.1-Base and introduces a hybrid inference architecture togglable via the chat template, along with post-training optimization specifically targeting tool use and multi-step agent tasks.
DeepSeek V3.1-Base is built upon the original V3 base checkpoint through a two-phase long-context extension approach:
Both phases use an expanded dataset of additional long documents compared to the original V3 training.
V3.1 introduces a hybrid thinking mode — one model supporting both thinking and non-thinking behavior by changing the chat template prefix:
</think> tag immediately, bypassing chain-of-thought<think> tags before answeringDeepSeek V3.1-Think is documented as reaching comparable answer quality to DeepSeek-R1-0528 while responding more quickly. API aliases: deepseek-chat → non-thinking mode; deepseek-reasoner → thinking mode.
Independent evaluations · Artificial Analysis
Evaluate specifications, pricing, and independent benchmark indices
| Model Details | |||
|---|---|---|---|
| General Info | |||
| Provider | DeepSeek | Anthropic | Anthropic |
| Release Date | August 21, 2025 | July 24, 2026 | June 9, 2026 |
| Knowledge Cutoff | Mar 2025 | May 2026 | - |
| Context & Limits | |||
| Context Window | 164K | 1M Best Context Window | 1M Best Context Window |
| Pricing (per 1M tokens) | |||
| Input Pricing | $0.56 Best Input Pricing | $5 | $10 |
| Output Pricing | $1.68 Best Output Pricing | $25 | $50 |
| Modalities | |||
| Inputs | text | textimage | textimagefile |
| Outputs | text | text | text |
| Benchmarks (0-100) | |||
| Intelligence Index | 21.4 | 63.1 Best Intelligence Index | 62.1 |
| Coding Index | - | 78.0 Best Coding Index | 76.5 |
| Agentic Index | - | 59.2 Best Agentic Index | 56.6 |
Graduate-level reasoning and expert Q&A evaluation.
Extremely difficult logical reasoning and knowledge.
Logical reasoning over long context windows.
Independent evaluation data provided by Artificial Analysis. To view the latest benchmarks and full details, visit their official site.