MiMo-V2.5-Pro is Xiaomi’s flagship model, delivering strong performance in general agentic capabilities, complex software engineering, and long-horizon tasks, with top rankings on benchmarks such as ClawEval, GDPVal, and SWE-bench Pro....
Tested on 5 benchmarks with 0.0% average. Top scores: Chatbot Arena Elo — Coding (1471.3%), Chatbot Arena Elo — Overall (1466.2%), Artificial Analysis — Coding Index (60.2%).
Chatbot Arena coding Elo. Human preference ranking specifically for coding tasks and technical questions.
Chatbot Arena overall Elo rating. Crowdsourced human preference ranking from blind head-to-head comparisons across all topics.
Artificial Analysis Coding Index. Composite coding quality score from multiple code benchmarks.
Artificial Analysis Quality Index. Composite quality score combining multiple benchmark results into a single metric.
Artificial Analysis Agentic Index. Composite score measuring agent capability across tool use and planning tasks.
- Typetext
- Context1.0M tokens (~524 books)
- ReleasedApr 2026
- LicenseOpen Source
- StatusActive
- Cost / Message~$0.002