BenchLeader
XiaomiOpen weightsReasoning modelNewFirst measured 22 Sept 2026

MiMo-V2.6-Pro

MiMo-V2.6-Pro is a Xiaomi open-weights reasoning model. It is not yet ranked: 4 independent results so far, and the index needs at least five across two categories. At $0.548 per million tokens blended it is mid-priced. Output speed of 125 tokens per second puts it faster than most, with a first answer in 18.4 s.

Blended price
$0.548/M
$0.440 in · $0.870 out
Output speed
125 tok/s
measured by Artificial Analysis
First answer
18 s
first token 3.02 s
Context
1.0M
Full answer
22 s
median, reasoning included
Cost per run
$0.133
one full Intelligence Index run
Released
date not published by our sources
How it scores by categoryDashed line = average model (50). One step of 15 = one standard deviation.
  1. Reasoning92
  2. Knowledge68
  3. Long context70
  4. Composite87

Versions

Xiaomi has shipped 2 models under this name. Each is ranked on its own results; a newer version often has fewer results so far, which holds its index nearer the average until more arrive.

ModelReleasedIndexRank
MiMo-V2.5-Pro23 Apr 202659.1#120
NewMiMo-V2.6-Prothis page

Benchmark results

One column per reasoning effort. Rank is among every configuration of every model on that benchmark. Hover a score for the run it came from.

Reasoning

Benchmarknot statedSource
Humanity's Last Exam (AA)not in index49.4%#15Artificial Analysis
CritPt26.6%#22Artificial Analysis

Coding

Benchmarknot statedSource
SciCode (AA)not in index60.9%#3Artificial Analysis

Agents & tools

Benchmarknot statedSource
GDPval (AA)not in index58.7%#8Artificial Analysis

Knowledge

Benchmarknot statedSource
AA-Omniscience8.4#69Artificial Analysis

Long context

Benchmarknot statedSource
AA-LCR86.3%#3Artificial Analysis

Composite

Benchmarknot statedSource
AA Intelligence Index46.3#18Artificial Analysis

Where it wins

Benchmarks where this configuration ranks in the top five of every configuration measured.

Where to run it

Every provider serving this model through OpenRouter, with throughput and first-token latency measured on live traffic over the last 30 minutes and each provider’s own price. Purple marks the best in each column.

ProviderSpeedFirst tokenInput $/MOutput $/MContextQuantisation
Xiaomi34 tok/s3.02 s$0.435$0.8701.0Mfp8

What a task costs

Estimates from list price, output speed and time to first answer for the best configuration. “With caching” assumes three-quarters of the input is served from the prompt cache. Reasoning tokens are not modelled.

WorkloadTokens in / outCostWith cachingTime
Chat reply400 / 300$0.000420.8 s
Summarise a 30-page report12,000 / 600$0.005823.2 s
Code edit6,000 / 1,500$0.003930.5 s
Agentic coding session60,000 / 4,000$0.03050.5 s
Structured extraction2,000 / 200$0.001120.0 s

See also

Data as of 22 Sept 2026. Compare with another model.

Cite as: BenchLeader, “MiMo-V2.6-Pro: benchmarks, pricing, speed and rank”, https://www.benchleader.com/models/mimo-v2-6-pro, data as of 22 Sept 2026.