DeepSeek R1 Distill Llama 70B by DeepSeek is a 70B dense model fine-tuned on DeepSeek R1 reasoning samples via SFT only, derived from Llama 3.3 70B Instruct.
Capabilities, design details, and architectural traits
DeepSeek R1 Distill Llama 70B is a 70B dense model produced by fine-tuning Llama 3.3 70B Instruct on reasoning samples generated by DeepSeek R1. It is the largest of the six distilled models released alongside DeepSeek R1, and the only one in the set derived from the Llama 3.3 family — chosen officially for its stronger reasoning capability compared to Llama 3.1 at the same parameter scale.
Derived from Llama 3.3 70B Instruct, originally licensed under the Llama 3.3 license. Released under the MIT License by DeepSeek, which permits commercial use, modifications, and derivative works including further distillation for training other LLMs. The Llama 3.3 base license terms also apply given the model's derivation.
Officially documented usage configurations for the R1 Distill series:
Independent evaluations · Artificial Analysis
Evaluate specifications, pricing, and independent benchmark indices
| Model Details | |||
|---|---|---|---|
| General Info | |||
| Provider | DeepSeek | Anthropic | Anthropic |
| Release Date | January 20, 2025 | July 24, 2026 | June 9, 2026 |
| Knowledge Cutoff | Jul 2024 | May 2026 | - |
| Context & Limits | |||
| Context Window | 131K | 1M Best Context Window | 1M Best Context Window |
| Pricing (per 1M tokens) | |||
| Input Pricing | $0.70 Best Input Pricing | $5 | $10 |
| Output Pricing | $1.10 Best Output Pricing | $25 | $50 |
| Modalities | |||
| Inputs | text | textimage | textimagefile |
| Outputs | text | text | text |
| Benchmarks (0-100) | |||
| Intelligence Index | 9.8 | 63.1 Best Intelligence Index | 62.1 |
| Coding Index | - | 78.0 Best Coding Index | 76.5 |
| Agentic Index | - | 59.2 Best Agentic Index | 56.6 |
Graduate-level reasoning and expert Q&A evaluation.
Extremely difficult logical reasoning and knowledge.
Logical reasoning over long context windows.
Independent evaluation data provided by Artificial Analysis. To view the latest benchmarks and full details, visit their official site.