MiMo-V2-Pro is Xiaomi's flagship foundation model, featuring over 1T total parameters and a 1M context length, deeply optimized for agentic scenarios. It is highly adaptable to general agent frameworks like...
Tested on 13 benchmarks · BenchGecko score 59.6. Top scores: Chatbot Arena Elo — Overall (1445.0%), Chatbot Arena Elo — Coding (1433.4%), LiveBench — Mathematics (77.0%).
DeepSeek V4 Flash 0731 scores 60.1 (101% as good) at $0.02/1M input · 98% cheaper
Regularly refreshed coding problems that avoid data contamination. New problems added monthly to prevent memorization.
LiveBench coding tasks that require multi-step reasoning and tool use. Tests planning and execution of complex coding workflows.
Regularly refreshed reasoning problems testing logical deduction, spatial reasoning, and analytical thinking.
Fresh data analysis tasks testing ability to interpret tables, charts, and statistical data.
Regularly updated math problems that test numerical reasoning, algebra, calculus, and combinatorics.
- Typetext
- Context1.0M tokens (~524 books)
- ReleasedMar 2026
- LicenseProprietary
- StatusActive
- Cost / Message~$0.005
Frequently Asked Questions
Related Models
Qwen2.5 72B Instruct · Alibaba QwenDeepSeek V4 Flash 0731 · DeepSeekStable Beluga 2 · UnknownQwen-14B · Alibaba QwenClaude Opus 4.5 · AnthropicKey facts · as of 2026-05-03
- MiMo-V2-Pro by xiaomi. BenchGecko score 59.6, rank 90 of 312 scored models (normalized average of public benchmark scores).
- List price $1.00 input · $3.00 output per 1M tokens (as of 2026-05-03).
How to cite · data as of 2026-05-03
MiMo-V2-Pro · benchmarks, pricing and providers. BenchGecko, data as of 2026-05-03. https://benchgecko.ai/model/mimo-v2-pro
Credit "Source: BenchGecko" with a link. Prices per provider and Gecko Tests are BenchGecko data (CC BY 4.0); benchmark scores keep their original source, listed in the JSON. JSON · llms.txt · MCP