MiMo V2.5 with Xiaomi thinking enabled. It supports deep reasoning, tool calling, structured outputs, and web search with up to 1M context.
Added Jun 3, 2026
Model weightsContext Window
1.0M
Max Output
131.1K
Avg output tokens (7d)
1.3K tokens
Input Price (Auto)
$0.14/1M
Output Price (Auto)
$0.28/1M
Cache Read (Auto)
$0.0028/1M
Capabilities
Benchmarks
Benchmarks
Performance metrics and benchmarks
Sourced from Artificial Analysis.
Intelligence Index
25.2
Coding Index
56.8
Agentic Index
15.8
Agentic work
AutomationBench-AA
Workflow automation with guardrail penalties
18.4%
Better than 40% of models compared
AA-Briefcase
Agentic knowledge work (Elo)
752 Elo
Better than 33% of models compared
GDPval-AA v2
Economically valuable tasks (Elo)
986 Elo
Better than 46% of models compared
Document reasoning
AA-LCR v1.1
Long context reasoning with updated grading
73.0%
Better than 69% of models compared
Reasoning
HLE
Humanity's Last Exam
27.2%
Better than 74% of models compared
IFBench
Instruction-following benchmark
67.1%
Better than 80% of models compared
CritPt
Research-level physics reasoning
3.7%
Coding
Terminal-Bench v4.0
Practical coding and terminal tasks
0.0%
Better than 16% of models compared
SciCode
Python programming for scientific computing
43.9%
Better than 36% of models compared
Knowledge
AA-Omniscience Accuracy
Proportion of correctly answered questions
16.8%
AA-Omniscience Hallucination Rate
Rate of incorrect answers among non-correct responses
31.9%
Legacy benchmarks
GPQA Diamond (legacy)
Graduate-level scientific reasoning
84.9%
Better than 78% of models compared
Terminal-Bench Hard (legacy)
Agentic coding and terminal use
41.7%
Better than 88% of models compared
T²-Bench Telecom (legacy)
Conversational AI agents in dual-control scenarios
90.6%
Better than 85% of models compared
AA-LCR (unversioned / legacy)
Long context reasoning evaluation
73.0%
Better than 69% of models compared
GDPval-AA (unversioned / legacy)
Economically valuable tasks
24.3%
Last updated Oct 1, 2026
Artificial AnalysisProviders
Choose explicit providers for this model. Auto routing remains available as the default option.
Loading provider options…
Related text models
Compare MiMo V2.5 Thinking with similar models from the same provider or model family.
MiMo V2.6 Flash
xiaomi/mimo-v2.6-flashMiMo V2.6 Flash is Xiaomi's native omnimodal 309B-parameter mixture-of-experts model, activating 15B parameters per token. It balances intelligence, efficiency, and cost for coding, general agents, visual tasks, and cybersecurity, with text, image, video, and audio understanding and a 1M-token context window.
MiMo V2.6 Pro
xiaomi/mimo-v2.6-proMiMo V2.6 Pro is Xiaomi's flagship native omnimodal model, built with 1.02T total parameters and 42B active parameters per token. It is designed for demanding coding, long-horizon agent workflows, visual tasks, research, and cybersecurity, with text, image, video, and audio understanding and a 1M-token context window.
MiMo V2.6 Pro UltraSpeed
xiaomi/mimo-v2.6-pro-ultraspeedMiMo V2.6 Pro UltraSpeed is the latency-optimized serving mode for Xiaomi's flagship MiMo V2.6 Pro checkpoint. It preserves Pro-level quality while delivering up to 20x faster output for real-time coding, interactive agents, and other latency-sensitive workflows, with native text, image, video, and audio understanding and a 1M-token context window.
MiMo V2.5 Pro Thinking
xiaomi/mimo-v2.5-pro:thinkingMiMo V2.5 Pro with Xiaomi thinking enabled for coding, long-context reasoning, and agentic orchestration.
MiMo V2.5
xiaomi/mimo-v2.5MiMo V2.5 is Xiaomi's full-modal understanding model for agent workflows. It supports tool calling, structured outputs, and web search with up to 1M context.
MiMo V2.5 Pro
xiaomi/mimo-v2.5-proMiMo V2.5 Pro is Xiaomi's long-context flagship general model for coding and agentic orchestration. It supports tool calling and structured outputs with up to 1M context.