Toolbit.aiToolbit.ai

Find, compare, and explore the best AI tools to match your specific tasks and use cases.

Explore

  • AI Search
  • Compare ToolsNew
  • Browse Categories
  • Trending Tools
  • Most Popular
  • New Additions

Resources

  • Updates HubNew
  • AI News
  • ModelsNew
  • Blog Articles
  • NewsletterNew

Company

  • Launch a Tool
  • Advertise with Us
  • Guest Post
  • Contact Us
© 2026 Toolbit.ai. All rights reserved.
Privacy PolicyTerms & ConditionsDisclaimer
There's An AI For That favicon
There's An AI For That•The front page of AI for everyone
Toolbit.ai
Toolbit.ai
UpdatesNew
Blog
Toolbit.ai
Toolbit.ai
Toolbit.ai
Toolbit.ai
UpdatesNew
Blog
Sign in
  1. Home
  2. Updates
  3. Models
  4. MiMo-V2.5
Xiaomi
Released April 22, 2026

MiMo-V2.5

Xiaomi MiMo-V2.5 is an open-source Xiaomi model for agentic coding, long-horizon tasks, and 1M-token context with MoE and hybrid attention.

Visit XiaomiAnnouncement
Inputs
Text
Audio
Image
Video
Outputs
Text

Model Overview

Capabilities, design details, and architectural traits

MiMo-V2.5 - open-source agentic coding model

MiMo-V2.5 is built around long, structured work. It is presented as an open-source model for agentic coding and other long-horizon tasks that need sustained context.

DifferentiatorDocumented trait
1M-token contextSupports up to 1M tokens of context.
Mixture-of-Experts designUses a 1.02T-parameter MoE model with 42B active parameters.
Hybrid attention + MTPUses hybrid attention and Multi-Token Prediction from the MiMo-V2-Flash design.
Open-source releaseThe MiMo-V2.5 series is open-sourced under the MIT license.

Work pattern

The model is aimed at tasks that stay active across long spans of text and tool use. The documented focus is on agentic coding and long-horizon execution rather than a short chat exchange.

Series identity

MiMo-V2.5 belongs to a family that includes Pro, Omni, Flash, and TTS variants. The series is routed through Xiaomi MiMo with unified model naming across products.

Benchmark Performance

Independent evaluations · Artificial Analysis

38.0%
Intelligence
56.8%
Coding Index
24.4%
Agentic Index

Accuracy & Capability Details

GPQA - Graduate Science84.9%
Humanity's Last Exam27.2%
SciCode - Scientific Coding43.1%
Instruction Following67.1%
Long Context Reasoning68.3%
τ²-Bench - Agentic Tasks90.6%
TerminalBench - System Control41.7%
Specs
Context window
1.0Mtokens
Input pricing
$0.14per 1M tokens
Output pricing
$0.28per 1M tokens

Prices in USD.

Key Capabilities & Ratings

Long Context90%
Visual Reasoning80%
General Knowledge80%
Logical Logic80%
Multimodal Inputs80%

Compare Models Side-by-Side

Evaluate specifications, pricing, and independent benchmark indices

Model Details
General Info
ProviderXiaomiAnthropicAnthropic
Release DateApril 22, 2026July 24, 2026June 9, 2026
Knowledge Cutoff-May 2026-
Context & Limits
Context Window
1.0M
Best Context Window
1M1M
Pricing (per 1M tokens)
Input Pricing
$0.14
Best Input Pricing
$5$10
Output Pricing
$0.28
Best Output Pricing
$25$50
Modalities
Inputs
textaudioimagevideo
textimage
textimagefile
Outputs
text
text
text
Benchmarks (0-100)
Intelligence Index38.0
63.1
Best Intelligence Index
62.1
Coding Index56.8
78.0
Best Coding Index
76.5
Agentic Index24.4
59.2
Best Agentic Index
56.6
MiMo-V2.5
Claude Opus 5
Claude Fable 5

GPQA Benchmark

Graduate-level reasoning and expert Q&A evaluation.

85%
MiMo-V2.5
GPQA Benchmark
Score: 85%
MiMo-V2.5
93%
Claude Opus 5
GPQA Benchmark
Score: 93%
Claude Opus 5
93%
Claude Fable 5
GPQA Benchmark
Score: 93%
Claude Fable 5

Humanity's Last Exam

Extremely difficult logical reasoning and knowledge.

27%
MiMo-V2.5
Humanity's Last Exam
Score: 27%
MiMo-V2.5
55%
Claude Opus 5
Humanity's Last Exam
Score: 55%
Claude Opus 5
56%
Claude Fable 5
Humanity's Last Exam
Score: 56%
Claude Fable 5

Long Context Reasoning

Logical reasoning over long context windows.

68%
MiMo-V2.5
Long Context Reasoning
Score: 68%
MiMo-V2.5
76%
Claude Opus 5
Long Context Reasoning
Score: 76%
Claude Opus 5
77%
Claude Fable 5
Long Context Reasoning
Score: 77%
Claude Fable 5

Independent evaluation data provided by Artificial Analysis. To view the latest benchmarks and full details, visit their official site.

Back to all models