Quality
Highest overall scores
- 1
Xiaomi: MiMo-V2.5-Pro42.9
- 2
Xiaomi: MiMo-V2.538
Compare Xiaomi models across answer quality, coding, complex tasks, speed, price, and context size.
Highest overall scores
Fastest output in this sample
Lowest price per 1M tokens
Overall capability · Higher is better
Output tokens per second · Higher is better
USD per 1M tokens · Lower is better
Quality benchmarks
Performance and price
2 models in the current catalogue.
| Model | Links | |||||
|---|---|---|---|---|---|---|
| Xiaomi: MiMo-V2.5-Pro | 42.9 | 28 t/s | 1.1M | $0.43 | $0.87 | |
| Xiaomi: MiMo-V2.5 | 38 | 70 t/s | 1.1M | $0.14 | $0.28 |
Compare named evaluation records available for Xiaomi models. Missing results remain blank rather than being estimated.
| Benchmark | Area | MiMo-V2.5-Pro | MiMo-V2.5 | Source |
|---|---|---|---|---|
| Intelligence Index | overallQuality | 42.9 | 38 | OpenRouter / artificial-analysis |
| Coding Index | coding | 60.2 | 56.8 | OpenRouter / artificial-analysis |
| Agentic Index | agentic | 29.5 | 24.4 | OpenRouter / artificial-analysis |
| GPQA | GPQA Diamond | — | 84.9 | OpenRouter model benchmarks |
| Humanity's Last Exam | HLE | — | 27.2 | OpenRouter model benchmarks |
| IFBench | IFBench | — | 67.1 | OpenRouter model benchmarks |
| τ²-Bench Telecom | τ²-Bench Telecom | — | 90.6 | OpenRouter model benchmarks |
| AA-LCR | AA-LCR | — | 68.3 | OpenRouter model benchmarks |
| GDPval-AA | GDPval-AA | — | 32.5 | OpenRouter model benchmarks |
| CritPt | CritPt | — | 3.7 | OpenRouter model benchmarks |
| SciCode | SciCode | — | 43.1 | OpenRouter model benchmarks |
| Terminal-Bench Hard | Terminal-Bench Hard | — | 41.7 | OpenRouter model benchmarks |
| AA-Omniscience Accuracy | AA-Omniscience Accuracy | — | 16.8 | OpenRouter model benchmarks |
| AA-Omniscience Non-Hallucination Rate | AA-Omniscience Non-Hallucination Rate | — | 68.1 | OpenRouter model benchmarks |