MiMo V2.6 Pro UltraSpeed is the latency-optimized serving mode for Xiaomi's flagship MiMo V2.6 Pro checkpoint. It preserves Pro-level quality while delivering up to 20x faster output for real-time coding, interactive agents, and other latency-sensitive workflows, with native text, image, video, and audio understanding and a 1M-token context window.
Added Sep 21, 2026
Context Window
1.0M
Max Output
131.1K
Input Price (Auto)
$4.35/1M
Output Price (Auto)
$8.70/1M
Cache Read (Auto)
$0.036/1M
Capabilities
Benchmarks
Benchmarks
Performance metrics and benchmarks
Sourced from Artificial Analysis.
Intelligence Index
46.3
Agentic work
AutomationBench-AA
Workflow automation with guardrail penalties
58.6%
Better than 88% of models compared
AutomationBench-AA Tasks Completed
Fully completed workflows without guardrail violations
25.4%
Better than 42% of models compared
AA-Briefcase
Agentic knowledge work (Elo)
1522 Elo
Better than 89% of models compared
GDPval-AA v2
Economically valuable tasks (Elo)
1673 Elo
Better than 97% of models compared
Document reasoning
GDP.pdf
Professional PDF reasoning: all-pass rate
19.2%
Better than 73% of models compared
AA-LCR v1.1
Long context reasoning with updated grading
86.3%
Better than 99% of models compared
Reasoning
HLE
Humanity's Last Exam
49.4%
Better than 98% of models compared
Coding
Terminal-Bench v4.0
Practical coding and terminal tasks
34.8%
Better than 89% of models compared
SciCode
Python programming for scientific computing
60.9%
Better than 98% of models compared
Legacy benchmarks
AA-LCR (unversioned / legacy)
Long context reasoning evaluation
86.3%
Better than 99% of models compared
Last updated Sep 22, 2026
Artificial AnalysisProviders
Auto routing is available for this model. Explicit provider selection is not available.
Loading provider options…
Related text models
Compare MiMo V2.6 Pro UltraSpeed with similar models from the same provider or model family.
MiMo V2.6 Pro
xiaomi/mimo-v2.6-proMiMo V2.6 Pro is Xiaomi's flagship native omnimodal model, built with 1.02T total parameters and 42B active parameters per token. It is designed for demanding coding, long-horizon agent workflows, visual tasks, research, and cybersecurity, with text, image, video, and audio understanding and a 1M-token context window.
MiMo V2.5 Pro Thinking
xiaomi/mimo-v2.5-pro:thinkingMiMo V2.5 Pro with Xiaomi thinking enabled for coding, long-context reasoning, and agentic orchestration.
MiMo V2.5 Pro
xiaomi/mimo-v2.5-proMiMo V2.5 Pro is Xiaomi's long-context flagship general model for coding and agentic orchestration. It supports tool calling and structured outputs with up to 1M context.
MiMo V2.6 Flash
xiaomi/mimo-v2.6-flashMiMo V2.6 Flash is Xiaomi's native omnimodal 309B-parameter mixture-of-experts model, activating 15B parameters per token. It balances intelligence, efficiency, and cost for coding, general agents, visual tasks, and cybersecurity, with text, image, video, and audio understanding and a 1M-token context window.
MiMo V2.5 Thinking
xiaomi/mimo-v2.5:thinkingMiMo V2.5 with Xiaomi thinking enabled. It supports deep reasoning, tool calling, structured outputs, and web search with up to 1M context.
MiMo V2.5
xiaomi/mimo-v2.5MiMo V2.5 is Xiaomi's full-modal understanding model for agent workflows. It supports tool calling, structured outputs, and web search with up to 1M context.