DeepSeek R1 by DeepSeek is a 671B/37B-active MoE reasoning model trained via a two-stage RL and two-stage SFT pipeline with cold-start data. MIT licensed, open weights.
Capabilities, design details, and architectural traits
DeepSeek R1 is DeepSeek's first-generation dedicated reasoning model. It is built on DeepSeek V3-Base and trained through a documented four-stage pipeline combining two RL stages and two SFT stages — designed specifically to produce long chain-of-thought reasoning without the readability and consistency problems exhibited by its predecessor, DeepSeek R1-Zero.
The four-stage training pipeline is the defining characteristic of DeepSeek R1:
Independent evaluations · Artificial Analysis
Evaluate specifications, pricing, and independent benchmark indices
| Model Details | |||
|---|---|---|---|
| General Info | |||
| Provider | DeepSeek | Anthropic | Anthropic |
| Release Date | May 28, 2025 | July 24, 2026 | June 9, 2026 |
| Knowledge Cutoff | Jul 2024 | May 2026 | - |
| Context & Limits | |||
| Context Window | 164K | 1M Best Context Window | 1M Best Context Window |
| Pricing (per 1M tokens) | |||
| Input Pricing | $1.35 Best Input Pricing | $5 | $10 |
| Output Pricing | $3 Best Output Pricing | $25 | $50 |
| Modalities | |||
| Inputs | text | textimage | textimagefile |
| Outputs | text | text | text |
| Benchmarks (0-100) | |||
| Intelligence Index | 20.4 | 63.1 Best Intelligence Index | 62.1 |
| Coding Index | - | 78.0 Best Coding Index | 76.5 |
| Agentic Index | - | 59.2 Best Agentic Index | 56.6 |
Graduate-level reasoning and expert Q&A evaluation.
Extremely difficult logical reasoning and knowledge.
Logical reasoning over long context windows.
Independent evaluation data provided by Artificial Analysis. To view the latest benchmarks and full details, visit their official site.