DeepSeek V4 Flash by DeepSeek is a 284B/13B-active MoE model with hybrid CSA+HCA attention, three reasoning modes, and 1M-token context. Fast, efficient, MIT licensed.
Capabilities, design details, and architectural traits
DeepSeek V4 Flash is the efficiency-tier model in DeepSeek's V4 series: a 284B total parameter, 13B active parameter Mixture-of-Experts language model sharing the same hybrid attention architecture as DeepSeek V4 Pro, positioned as the fast and economical option within the V4 family. Both V4 models support a 1 million-token context window.
V4 Flash shares the three core architectural innovations introduced across the V4 series:
Three configurable reasoning effort modes per request, identical in structure to V4 Pro:
Independent evaluations · Artificial Analysis
Evaluate specifications, pricing, and independent benchmark indices
| Model Details | |||
|---|---|---|---|
| General Info | |||
| Provider | DeepSeek | Anthropic | Anthropic |
| Release Date | April 24, 2026 | July 24, 2026 | June 9, 2026 |
| Knowledge Cutoff | - | May 2026 | - |
| Context & Limits | |||
| Context Window | 1.0M Best Context Window | 1M | 1M |
| Pricing (per 1M tokens) | |||
| Input Pricing | $0.13 Best Input Pricing | $5 | $10 |
| Output Pricing | $0.28 Best Output Pricing | $25 | $50 |
| Modalities | |||
| Inputs | text | textimage | textimagefile |
| Outputs | text | text | text |
| Benchmarks (0-100) | |||
| Intelligence Index | 39.0 | 63.1 Best Intelligence Index | 62.1 |
| Coding Index | 52.0 | 78.0 Best Coding Index | 76.5 |
| Agentic Index | 30.3 | 59.2 Best Agentic Index | 56.6 |
Graduate-level reasoning and expert Q&A evaluation.
Extremely difficult logical reasoning and knowledge.
Logical reasoning over long context windows.
Independent evaluation data provided by Artificial Analysis. To view the latest benchmarks and full details, visit their official site.