Toolbit.aiToolbit.ai

Find, compare, and explore the best AI tools to match your specific tasks and use cases.

Explore

  • AI Search
  • Compare ToolsNew
  • Browse Categories
  • Trending Tools
  • Most Popular
  • New Additions

Resources

  • Updates HubNew
  • AI News
  • ModelsNew
  • Blog Articles
  • NewsletterNew

Company

  • Launch a Tool
  • Advertise with Us
  • Guest Post
  • Contact Us
© 2026 Toolbit.ai. All rights reserved.
Privacy PolicyTerms & ConditionsDisclaimer
There's An AI For That favicon
There's An AI For That•The front page of AI for everyone
Toolbit.ai
Toolbit.ai
UpdatesNew
Blog
Toolbit.ai
Toolbit.ai
Toolbit.ai
Toolbit.ai
UpdatesNew
Blog
Sign in
  1. Home
  2. Updates
  3. Models
  4. Mistral Small 3.1
Mistral
Released March 17, 2025Cutoff October 2023

Mistral Small 3.1

Mistral Small 3.1 is Mistral AI's vision upgrade to the text-only Small 3, adding image understanding, a Tekken tokenizer, and a 128k context window.

Visit MistralAnnouncement
Inputs
Text
Image
Outputs
Text

Model Overview

Capabilities, design details, and architectural traits

Mistral Small 3.1 - Vision Added to Small 3 Without Losing Text Performance

Mistral Small 3.1 builds on the text-only Mistral Small 3 (2501) by adding vision understanding and a longer context window, while stating that text performance is not compromised in the process.

TraitDocumented Behavior
Text-to-Vision UpgradeAdds vision understanding on top of the previously text-only Mistral Small 3, without reducing text performance.
Tekken TokenizerUses a Tekken tokenizer with a 131k-token vocabulary.
128k Context ExtensionExpands the context window up to 128k tokens, longer than the prior Small 3 release.

Released as Base and Instruct Together

Mistral AI released both Small 3.1 Base and Instruct checkpoints at the same time, explicitly to let the community build downstream reasoning models on top of the base weights rather than only shipping a closed instruction-tuned version.

Benchmark Performance

Independent evaluations · Artificial Analysis

14.9%
Intelligence
26.3%
Coding Index
5.3%
Agentic Index

Accuracy & Capability Details

GPQA - Graduate Science45.4%
Humanity's Last Exam4.3%
SciCode - Scientific Coding26.5%
Instruction Following29.9%
Long Context Reasoning22.0%
τ²-Bench - Agentic Tasks25.1%
TerminalBench - System Control7.6%
Specs
Context window
128Ktokens
Input pricing
$0.10per 1M tokens
Output pricing
$0.30per 1M tokens

Prices in USD.

Key Capabilities & Ratings

Coding90%
Advanced Math70%
Legal & E-Discovery70%
Financial Analysis70%
Translation & Lang70%

Compare Models Side-by-Side

Evaluate specifications, pricing, and independent benchmark indices

Model Details
General Info
ProviderMistralAnthropicAnthropic
Release DateMarch 17, 2025July 24, 2026June 9, 2026
Knowledge CutoffOct 2023May 2026-
Context & Limits
Context Window128K
1M
Best Context Window
1M
Best Context Window
Pricing (per 1M tokens)
Input Pricing
$0.10
Best Input Pricing
$5$10
Output Pricing
$0.30
Best Output Pricing
$25$50
Modalities
Inputs
textimage
textimage
textimagefile
Outputs
text
text
text
Benchmarks (0-100)
Intelligence Index14.9
63.1
Best Intelligence Index
62.1
Coding Index26.3
78.0
Best Coding Index
76.5
Agentic Index5.3
59.2
Best Agentic Index
56.6
Mistral Small 3.1
Claude Opus 5
Claude Fable 5

GPQA Benchmark

Graduate-level reasoning and expert Q&A evaluation.

45%
Mistral Small 3.1
GPQA Benchmark
Score: 45%
Mistral Small 3.1
93%
Claude Opus 5
GPQA Benchmark
Score: 93%
Claude Opus 5
93%
Claude Fable 5
GPQA Benchmark
Score: 93%
Claude Fable 5

Humanity's Last Exam

Extremely difficult logical reasoning and knowledge.

4%
Mistral Small 3.1
Humanity's Last Exam
Score: 4%
Mistral Small 3.1
55%
Claude Opus 5
Humanity's Last Exam
Score: 55%
Claude Opus 5
56%
Claude Fable 5
Humanity's Last Exam
Score: 56%
Claude Fable 5

Long Context Reasoning

Logical reasoning over long context windows.

22%
Mistral Small 3.1
Long Context Reasoning
Score: 22%
Mistral Small 3.1
76%
Claude Opus 5
Long Context Reasoning
Score: 76%
Claude Opus 5
77%
Claude Fable 5
Long Context Reasoning
Score: 77%
Claude Fable 5

Independent evaluation data provided by Artificial Analysis. To view the latest benchmarks and full details, visit their official site.

Back to all models