MiMo-V2.5
The 5th most-used model on the router over the last 7 days, with 4.2% of tokens. Weekly share peaked at 18.0% in the week of Jul 20, 2026, with 83T tokens processed since May 4, 2026.
- Launch
- Apr 22, 2026
- Context
- 1M tokens
- Modality
- text, image, audio and video → text
- Reasoning
- yes
- License
- Open weights
- Input
- $0.14 / 1M
- Output
- $0.28 / 1M
Weekly token share
Share of the router's weekly volume, since the first week with volume
See the numbers
| Week | Share |
|---|---|
| Aug 31, 2026 | 2.037% |
| Aug 24, 2026 | 8.086% |
| Aug 17, 2026 | 10.644% |
| Aug 10, 2026 | 5.033% |
| Aug 3, 2026 | 7.811% |
| Jul 27, 2026 | 11.102% |
| Jul 20, 2026 | 18.027% |
| Jul 13, 2026 | 14.846% |
| Jul 6, 2026 | 11.302% |
| Jun 29, 2026 | 9.377% |
| Jun 22, 2026 | 9.593% |
| Jun 15, 2026 | 8.449% |
| Jun 8, 2026 | 8.048% |
| Jun 1, 2026 | 6.057% |
| May 25, 2026 | 4.598% |
| May 18, 2026 | 0.326% |
| May 11, 2026 | 0.049% |
| May 4, 2026 | 0.039% |
Debuted in the week of May 4, 2026 and peaked at 18.0% 11 weeks later. The latest full week came in at 2.0%, 11% of peak.
What this model is used for new
Tasks where the model ranks among the leaders, and its share of each
Plus 5 more tasks with a smaller share.
Ranks among the leaders in 17 of the 29 classified tasks. Its largest share is in memory extraction, 12.0% of the task's tokens; in the heaviest task where it appears, workflow execution (24.6% of classified volume), it holds 3.3%.
Snapshot of the rolling 7-day window through Sep 10, 2026. Share: the model's tokens in the task over the task's total tokens. Weight: the task's share of classified volume.
Where to run it
Endpoints in the source's catalog, from cheapest to most expensive on input
| Provider | Input | Output | Context | Quantization | Zero retention |
|---|---|---|---|---|---|
| GMICloud | $0.119 | $0.238 | 1M | fp8 | no |
| DeepInfra | $0.133 | $0.266 | 256k | fp8 | yes |
| Xiaomi | $0.14 | $0.28 | 1M | fp8 | no |
| StreamLake | $0.168 | $0.336 | 1M | not reported | no |
| Novita | $0.168 | $0.336 | 1M | fp8 | yes |
5 endpoints across 5 providers, with input from $0.119 to $0.168 per 1M tokens. 2 offer zero data retention.
Prices in US dollars per million tokens, from the latest archived endpoint catalog.
Benchmarks
Artificial Analysis indexes and benchmarks run by the source, with cost per task
GPQA Diamond 76.3%
$0.0099 per task, 1353 tasks. 82nd of 131 evaluated. Median of evaluated models: 80.2%, $0.024 per task.
τ-bench verified, airline 70.9%
$0.0081 per task, 400 tasks. 59th of 122 evaluated. Median of evaluated models: 70.7%, $0.056 per task.
Artificial Analysis indexes, exposed by the catalog; benchmarks run by the source and archived on Sep 11, 2026. The gray mark is the median of evaluated models.
Against the week's largest models
The five largest proprietary and five largest open-weights models over the last 7 days
| Model | Lab | 7-day share | Blended price | Context | Intelligence | Weeks in top 10 |
|---|---|---|---|---|---|---|
| Proprietary | ||||||
| GPT-5.6 LunaOpenAI | OpenAI | 11.5% | $0.45 | 1M | 37.5 | 6 |
| Gemini 3.8 FlashGoogle | 2.1% | $1.50 | 1M | 41.2 | 0 | |
| Muse Spark 1.3 ContributorMeta | Meta | 1.5% | $0.125 | 1M | — | 0 |
| Claude Opus 5Anthropic | Anthropic | 1.4% | $10.00 | 1M | 50.7 | 1 |
| GPT-5.6 SolOpenAI | OpenAI | 1.4% | $4.00 | 1M | 47.1 | 0 |
| Open weights | ||||||
| Hy4 previewTencent | Tencent | 15.5% | $1.25 | 1M | — | 1 |
| DeepSeek V4 Flash 0731DeepSeek | DeepSeek | 10.0% | $0.094 | 1.31M | 34.5 | 5 |
| GLM 5.3 FlashZ.ai (GLM) | Z.ai (GLM) | 10.0% | $0.237 | 1.31M | 41.9 | 2 |
| MiMo-V2.5this pageXiaomi | Xiaomi | 4.2% | $0.175 | 1M | 22.3 | 15 |
| DeepSeek V4 Flash 0423DeepSeek | DeepSeek | 3.9% | $0.111 | 1M | 24.8 | 19 |
Share over the last 7 days (Sep 4, 2026 to Sep 10, 2026). Blended price in US dollars per million tokens, weighted 75% input and 25% output. A dash means the number is not published, not zero.