Toolbit.aiToolbit.ai

Find, compare, and explore the best AI tools to match your specific tasks and use cases.

Explore

  • AI Search
  • Compare ToolsNew
  • Browse Categories
  • Trending Tools
  • Most Popular
  • New Additions

Resources

  • Updates HubNew
  • AI News
  • ModelsNew
  • Blog Articles
  • NewsletterNew

Company

  • Launch a Tool
  • Advertise with Us
  • Guest Post
  • Contact Us
© 2026 Toolbit.ai. All rights reserved.
Privacy PolicyTerms & ConditionsDisclaimer
There's An AI For That favicon
There's An AI For That•The front page of AI for everyone
Toolbit.ai
Toolbit.ai
UpdatesNew
Blog
Toolbit.ai
Toolbit.ai
Toolbit.ai
Toolbit.ai
UpdatesNew
Blog
Sign in
  1. Home
  2. Updates
  3. Models
  4. Mistral Large
Mistral
Released February 26, 2024Cutoff November 2024

Mistral Large

Mistral Large by Mistral AI is a 123B flagship model with 128k context, 80+ coding languages, native function calling, and single-node inference design.

Visit MistralAnnouncement
Inputs
Text
File
Outputs
Text

Model Overview

Capabilities, design details, and architectural traits

Mistral Large – Flagship Single-Node Dense Model

Mistral Large is Mistral AI's top-tier commercial model line, positioned for high-complexity reasoning, multilingual tasks, and long-context applications. Its defining engineering constraint is single-node inference suitability: at 123B parameters (Large 2), it is explicitly sized to run at large throughput on a single node — a deliberate departure from the multi-node requirements typical of frontier-scale models.

Architecture & Training Focus

Large 2 was trained on a heavily code-weighted corpus following Mistral's experience with Codestral. A specific training objective was response conciseness — Mistral officially documented and benchmarked output length against Claude 3 Opus, Claude 3.5 Sonnet, Llama 3.1, and GPT-4o, targeting shorter, production-suitable responses.

Version Lineage (Official)

  • Large (2402): First release; 32k context, native function calling, Azure launch partner
  • Large 2 (2407): 123B, 128k context, code-heavy training, hallucination-reduction focus
  • Large 2.1 (2411): Top-tier high-complexity release, November 2024
  • Large 3 (2512): Sparse MoE, 41B active / 675B total parameters, Apache 2.0

Benchmark Performance

Independent evaluations · Artificial Analysis

4.1%
Intelligence

Accuracy & Capability Details

GPQA - Graduate Science35.1%
Humanity's Last Exam3.5%
SciCode - Scientific Coding20.8%
Specs
Context window
128Ktokens
Input pricing
$4per 1M tokens
Output pricing
$12per 1M tokens

Prices in USD.

Key Capabilities & Ratings

Image to text90%
Visual Reasoning80%
Logical Logic80%
Multimodal Inputs80%
Advanced Math70%

Compare Models Side-by-Side

Evaluate specifications, pricing, and independent benchmark indices

Model Details
General Info
ProviderMistralAnthropicAnthropic
Release DateFebruary 26, 2024July 24, 2026June 9, 2026
Knowledge CutoffNov 2024May 2026-
Context & Limits
Context Window128K
1M
Best Context Window
1M
Best Context Window
Pricing (per 1M tokens)
Input Pricing
$4
Best Input Pricing
$5$10
Output Pricing
$12
Best Output Pricing
$25$50
Modalities
Inputs
textfile
textimage
textimagefile
Outputs
text
text
text
Benchmarks (0-100)
Intelligence Index4.1
63.1
Best Intelligence Index
62.1
Coding Index-
78.0
Best Coding Index
76.5
Agentic Index-
59.2
Best Agentic Index
56.6
Mistral Large
Claude Opus 5
Claude Fable 5

GPQA Benchmark

Graduate-level reasoning and expert Q&A evaluation.

35%
Mistral Large
GPQA Benchmark
Score: 35%
Mistral Large
93%
Claude Opus 5
GPQA Benchmark
Score: 93%
Claude Opus 5
93%
Claude Fable 5
GPQA Benchmark
Score: 93%
Claude Fable 5

Humanity's Last Exam

Extremely difficult logical reasoning and knowledge.

4%
Mistral Large
Humanity's Last Exam
Score: 4%
Mistral Large
55%
Claude Opus 5
Humanity's Last Exam
Score: 55%
Claude Opus 5
56%
Claude Fable 5
Humanity's Last Exam
Score: 56%
Claude Fable 5

SciCode Benchmark

Scientific coding and mathematical modeling.

21%
Mistral Large
SciCode Benchmark
Score: 21%
Mistral Large
56%
Claude Opus 5
SciCode Benchmark
Score: 56%
Claude Opus 5
60%
Claude Fable 5
SciCode Benchmark
Score: 60%
Claude Fable 5

Independent evaluation data provided by Artificial Analysis. To view the latest benchmarks and full details, visit their official site.

Back to all models