Mistral Small 3 by Mistral AI is a 24B latency-optimized model with 150 tokens/s, 81% MMLU, local deployment on a single RTX 4090.
Capabilities, design details, and architectural traits
Mistral Small 3 is a 24B-parameter model explicitly engineered around forward-pass speed, using far fewer layers than competing models to minimize time per forward pass. Released under Apache 2.0, it targets the '80% of generative AI tasks' that demand robust instruction-following at very low latency.
Mistral explicitly notes the model is not trained with reinforcement learning or synthetic data, placing it earlier in the production pipeline than models like DeepSeek R1. This design choice makes it a clean base for community fine-tuning toward specialized reasoning capacities rather than a finished reasoning model.
Available as mistral-small-2501 on La Plateforme API and distributed across Hugging Face, Ollama, Kaggle, Together AI, Fireworks AI, and IBM WatsonX. Both pretrained base and instruction-tuned checkpoints are released.
Independent evaluations · Artificial Analysis
Evaluate specifications, pricing, and independent benchmark indices
| Model Details | |||
|---|---|---|---|
| General Info | |||
| Provider | Mistral | Anthropic | Anthropic |
| Release Date | January 30, 2025 | July 24, 2026 | June 9, 2026 |
| Knowledge Cutoff | Oct 2023 | May 2026 | - |
| Context & Limits | |||
| Context Window | 33K | 1M Best Context Window | 1M Best Context Window |
| Pricing (per 1M tokens) | |||
| Input Pricing | $0.10 Best Input Pricing | $5 | $10 |
| Output Pricing | $0.30 Best Output Pricing | $25 | $50 |
| Modalities | |||
| Inputs | text | textimage | textimagefile |
| Outputs | text | text | text |
| Benchmarks (0-100) | |||
| Intelligence Index | 6.7 | 63.1 Best Intelligence Index | 62.1 |
| Coding Index | - | 78.0 Best Coding Index | 76.5 |
| Agentic Index | - | 59.2 Best Agentic Index | 56.6 |
Graduate-level reasoning and expert Q&A evaluation.
Extremely difficult logical reasoning and knowledge.
Logical reasoning over long context windows.
Independent evaluation data provided by Artificial Analysis. To view the latest benchmarks and full details, visit their official site.