DeepSeek V3.2 by DeepSeek is a 671B/37B-active MoE model with DeepSeek Sparse Attention, scalable RL, and the first open-source integration of thinking with tool-use.
Capabilities, design details, and architectural traits
DeepSeek V3.2 is DeepSeek's flagship open-weight model built on three documented technical breakthroughs: DeepSeek Sparse Attention (DSA) for long-context efficiency, a scalable reinforcement learning framework for post-training compute scaling, and a large-scale agentic task synthesis pipeline that integrates reasoning directly into tool-use — the first open-source model to do so.
DeepSeek V3.2 is a 671B total parameter, 37B active parameter Mixture-of-Experts model. The MoE layer uses 256 expert networks per layer (up from 160 in V2), activating 8 per token: 1–2 shared experts handling common patterns plus 6–7 routed experts. Built on the same model structure as DeepSeek-V3.2-Exp.
DeepSeek V3.2 is documented as DeepSeek's first model to integrate thinking directly into tool-use, supporting two parallel modes:
A new developer role is introduced in the chat template, dedicated exclusively to search agent scenarios
Independent evaluations · Artificial Analysis
Evaluate specifications, pricing, and independent benchmark indices
| Model Details | |||
|---|---|---|---|
| General Info | |||
| Provider | DeepSeek | Anthropic | Anthropic |
| Release Date | December 1, 2025 | July 24, 2026 | June 9, 2026 |
| Knowledge Cutoff | - | May 2026 | - |
| Context & Limits | |||
| Context Window | 131K | 1M Best Context Window | 1M Best Context Window |
| Pricing (per 1M tokens) | |||
| Input Pricing | $0.28 Best Input Pricing | $5 | $10 |
| Output Pricing | $0.42 Best Output Pricing | $25 | $50 |
| Modalities | |||
| Inputs | text | textimage | textimagefile |
| Outputs | text | text | text |
| Benchmarks (0-100) | |||
| Intelligence Index | 25.1 | 63.1 Best Intelligence Index | 62.1 |
| Coding Index | - | 78.0 Best Coding Index | 76.5 |
| Agentic Index | - | 59.2 Best Agentic Index | 56.6 |
Graduate-level reasoning and expert Q&A evaluation.
Extremely difficult logical reasoning and knowledge.
Logical reasoning over long context windows.
Independent evaluation data provided by Artificial Analysis. To view the latest benchmarks and full details, visit their official site.