MiMo V2.5 Pro with Xiaomi thinking enabled for coding, long-context reasoning, and agentic orchestration.
Added Jun 3, 2026
Model weightsContext Window
1.0M
Max Output
131.1K
Avg output tokens (7d)
1.3K tokens
Input Price (Auto)
$0.43/1M
Output Price (Auto)
$0.87/1M
Cache Read (Auto)
$0.0036/1M
Capabilities
Benchmarks
Benchmarks
Performance metrics and benchmarks
Sourced from Artificial Analysis.
Intelligence Index
26.4
Coding Index
60.2
Agentic Index
22.7
Agentic work
AutomationBench-AA
Workflow automation with guardrail penalties
13.7%
Better than 41% of models compared
AA-Briefcase
Agentic knowledge work (Elo)
883 Elo
Better than 47% of models compared
GDPval-AA v2
Economically valuable tasks (Elo)
1186 Elo
Better than 63% of models compared
Document reasoning
GDP.pdf
Professional PDF reasoning: all-pass rate
4.0%
Better than 28% of models compared
AA-LCR v1.1
Long context reasoning with updated grading
79.7%
Better than 88% of models compared
Reasoning
HLE
Humanity's Last Exam
35.7%
Better than 87% of models compared
IFBench
Instruction-following benchmark
79.9%
Better than 98% of models compared
CritPt
Research-level physics reasoning
4.0%
Coding
Terminal-Bench v4.0
Practical coding and terminal tasks
0.0%
Better than 17% of models compared
SciCode
Python programming for scientific computing
50.6%
Better than 61% of models compared
Knowledge
AA-Omniscience Accuracy
Proportion of correctly answered questions
22.4%
AA-Omniscience Hallucination Rate
Rate of incorrect answers among non-correct responses
24.7%
Legacy benchmarks
GPQA Diamond (legacy)
Graduate-level scientific reasoning
86.6%
Better than 81% of models compared
Terminal-Bench Hard (legacy)
Agentic coding and terminal use
43.2%
Better than 90% of models compared
T²-Bench Telecom (legacy)
Conversational AI agents in dual-control scenarios
94.2%
Better than 92% of models compared
AA-LCR (unversioned / legacy)
Long context reasoning evaluation
79.7%
Better than 88% of models compared
GDPval-AA (unversioned / legacy)
Economically valuable tasks
34.3%
Last updated Sep 12, 2026
Artificial AnalysisProviders
Choose explicit providers for this model. Auto routing remains available as the default option.
Loading provider options…
Related text models
Compare MiMo V2.5 Pro Thinking with similar models from the same provider or model family.
MiMo V2.5 Pro (Crof)
xiaomi/mimo-v2.5-pro-crofMiMo V2.5 Pro is Xiaomi's long-context flagship general model for coding and agentic orchestration. This separately served variant is intended for users concerned about censorship on the regular Xiaomi MiMo V2.5 Pro.
MiMo V2.5 Pro Thinking (Crof)
xiaomi/mimo-v2.5-pro-crof:thinkingMiMo V2.5 Pro with Xiaomi thinking enabled for coding, long-context reasoning, and agentic orchestration. This separately served thinking variant is intended for users concerned about censorship on the regular Xiaomi MiMo V2.5 Pro.
MiMo V2.5 Pro
xiaomi/mimo-v2.5-proMiMo V2.5 Pro is Xiaomi's long-context flagship general model for coding and agentic orchestration. It supports tool calling and structured outputs with up to 1M context.
MiMo V2.5 Thinking
xiaomi/mimo-v2.5:thinkingMiMo V2.5 with Xiaomi thinking enabled. It supports deep reasoning, tool calling, structured outputs, and web search with up to 1M context.
MiMo V2.5
xiaomi/mimo-v2.5MiMo V2.5 is Xiaomi's full-modal understanding model for agent workflows. It supports tool calling, structured outputs, and web search with up to 1M context.
Synth 2.5 Pro Preview
synth-2.5-proSynth 2.5 Pro Preview is a low-cost text model designed for role-play, character dialogue, and collaborative storytelling.