MiMo V2.5 Pro with Xiaomi thinking enabled for coding, long-context reasoning, and agentic orchestration.
Added Jun 3, 2026
Model weightsContext Window
1.0M
Max Output
131.1K
Avg output tokens (7d)
632 tokens
Input Price (Auto)
$0.43/1M
Output Price (Auto)
$0.87/1M
Cache Read (Auto)
$0.0036/1M
Capabilities
Benchmarks
Benchmarks
Performance metrics and benchmarks
Sourced from Artificial Analysis.
Intelligence Index
26.0
Coding Index
60.2
Agentic Index
21.3
Agentic work
AutomationBench-AA
Workflow automation with guardrail penalties
13.7%
Better than 37% of models compared
Harvey LAB-AA
Legal agentic work criterion pass rate
73.3%
Better than 20% of models compared
AA-Briefcase
Agentic knowledge work (Elo)
881 Elo
Better than 40% of models compared
GDPval-AA v2
Economically valuable tasks (Elo)
1107 Elo
Better than 55% of models compared
Document reasoning
GDP.pdf
Professional PDF reasoning: all-pass rate
4.0%
Better than 22% of models compared
AA-LCR v1.1
Long context reasoning with updated grading
79.7%
Better than 85% of models compared
MLCR-AA
Medical long-context reasoning
9.4%
Better than 29% of models compared
Reasoning
HLE
Humanity's Last Exam
35.7%
Better than 84% of models compared
IFBench
Instruction-following benchmark
79.9%
Better than 98% of models compared
CritPt
Research-level physics reasoning
4.0%
Coding
Terminal-Bench v4.0
Practical coding and terminal tasks
0.0%
Better than 16% of models compared
SciCode
Python programming for scientific computing
50.6%
Better than 57% of models compared
Knowledge
AA-Omniscience Accuracy
Proportion of correctly answered questions
22.4%
AA-Omniscience Hallucination Rate
Rate of incorrect answers among non-correct responses
24.7%
Legacy benchmarks
GPQA Diamond (legacy)
Graduate-level scientific reasoning
86.6%
Better than 81% of models compared
Terminal-Bench Hard (legacy)
Agentic coding and terminal use
43.2%
Better than 90% of models compared
T²-Bench Telecom (legacy)
Conversational AI agents in dual-control scenarios
94.2%
Better than 92% of models compared
AA-LCR (unversioned / legacy)
Long context reasoning evaluation
79.7%
Better than 85% of models compared
GDPval-AA (unversioned / legacy)
Economically valuable tasks
30.4%
Last updated Oct 2, 2026
Artificial AnalysisProviders
Choose explicit providers for this model. Auto routing remains available as the default option.
Loading provider options…
Related text models
Compare MiMo V2.5 Pro Thinking with similar models from the same provider or model family.
MiMo V2.6 Pro
xiaomi/mimo-v2.6-proMiMo V2.6 Pro is Xiaomi's flagship native omnimodal model, built with 1.02T total parameters and 42B active parameters per token. It is designed for demanding coding, long-horizon agent workflows, visual tasks, research, and cybersecurity, with text, image, video, and audio understanding and a 1M-token context window.
MiMo V2.6 Pro UltraSpeed
xiaomi/mimo-v2.6-pro-ultraspeedMiMo V2.6 Pro UltraSpeed is the latency-optimized serving mode for Xiaomi's flagship MiMo V2.6 Pro checkpoint. It preserves Pro-level quality while delivering up to 20x faster output for real-time coding, interactive agents, and other latency-sensitive workflows, with native text, image, video, and audio understanding and a 1M-token context window.
MiMo V2.5 Pro
xiaomi/mimo-v2.5-proMiMo V2.5 Pro is Xiaomi's long-context flagship general model for coding and agentic orchestration. It supports tool calling and structured outputs with up to 1M context.
MiMo V2.6 Flash
xiaomi/mimo-v2.6-flashMiMo V2.6 Flash is Xiaomi's native omnimodal 309B-parameter mixture-of-experts model, activating 15B parameters per token. It balances intelligence, efficiency, and cost for coding, general agents, visual tasks, and cybersecurity, with text, image, video, and audio understanding and a 1M-token context window.
MiMo V2.5 Thinking
xiaomi/mimo-v2.5:thinkingMiMo V2.5 with Xiaomi thinking enabled. It supports deep reasoning, tool calling, structured outputs, and web search with up to 1M context.
MiMo V2.5
xiaomi/mimo-v2.5MiMo V2.5 is Xiaomi's full-modal understanding model for agent workflows. It supports tool calling, structured outputs, and web search with up to 1M context.