Quick answer: MiMo V2.5 Pro is Xiaomi's July 2025 reasoning model scoring 78.9% SWE-Bench Verified, 99.6% GSM8K, 86.2% MATH, 77.9% MMMU-Pro, and 34.0% HLE. Apache 2.0 — Xiaomi's open-source reasoning model with near-perfect math.
Where MiMo V2.5 Pro leads
Where it lags
Best for: Open-source math and reasoning pipelines; SWE-Bench class coding; Apache 2.0 deployments needing balanced math+coding.
MiMo V2.5 Pro is Xiaomi's July 2025 reasoning model from the MiMo (Mini Model) series — Xiaomi's AI research team's open-source model line. The V2.5 Pro is designed with a strong math-reasoning focus while also achieving competitive SWE-Bench scores.
The 99.6% GSM8K is near-perfect and among the highest reported on this benchmark. The 86.2% MATH is also excellent. The model appears to use chain-of-thought reasoning training similar to DeepSeek-R1, focusing on mathematical and code reasoning.
| Field | Value |
|---|---|
| Organization | Xiaomi |
| License | Apache 2.0 |
| Release date | July 2025 |
| Modality | Text only |
Open weights under Apache 2.0 — self-host at no cost.
| Benchmark | Score | Source | Date |
|---|---|---|---|
| SWE-Bench Verified | 78.9% | Benchgen evaluation | 2025-07 |
| GSM8K | 99.6% | Benchgen evaluation | 2025-07 |
| MATH | 86.2% | Benchgen evaluation | 2025-07 |
| MMMU-Pro | 77.9% | Benchgen evaluation | 2025-07 |
| Humanity's Last Exam | 34.0% | Benchgen evaluation | 2025-07 |
| MMLU-Pro | 68.5% | Benchgen evaluation | 2025-07 |
| Model | SWE-Bench | MATH | GSM8K | License |
|---|---|---|---|---|
| MiMo V2.5 Pro | 78.9% | 86.2% | 99.6% | Apache 2.0 |
| Kimi K2.6 | 80.2% | — | — | Apache 2.0 |
| DeepSeek-R1 0528 | — | 97.3% | 96.2% | MIT |
| Gemini 3 Flash | 78% | 97.5% | 96.8% | Proprietary |
MiMo V2.5 Pro leads on GSM8K (99.6%) among Apache 2.0 models. SWE-Bench is competitive (78.9% — near Kimi K2.6's 80.2%). For open-source math-heavy pipelines with SWE-Bench performance: MiMo V2.5 Pro is strong.
Specs from Xiaomi's MiMo V2.5 Pro release (July 2025) and Benchgen evaluations. Last updated 2026-07-24.
This model isn’t on any benchmark leaderboard yet.