BenchLeader

MiMo-V2.6-Pro vs Step 5 Preview

Verdict
  • MiMo-V2.6-Pro and Step 5 Preview are level on quality (64.3 vs 64.4).
  • MiMo-V2.6-Pro is stronger in agents & tools, composite, reasoning.
  • Step 5 Preview is stronger in knowledge, long context, coding, multimodal.
  • MiMo-V2.6-Pro is 2.6× cheaper ($0.548 vs $1.43 per 1M blended).
  • Step 5 Preview streams 1.5× faster (83 vs 54 tokens per second).
MetricMiMo-V2.6-ProStep 5 Preview
BenchLeader Index64.364.4
Agents & tools score55.1
Composite score84.881.7
Knowledge score66.870.5
Long context score69.170.1
Reasoning score88.878.6
Coding score68.8
Multimodal score60.0
Blended price $/M$0.548$1.43
Output speed54 tok/s83 tok/s
Time to first answer39.8 s27.7 s
Context window1.0M1M
SciCode58.9%
AA Intelligence Index46.343.7
AA-LCR86.3%88.3%
MMMU-Pro76.4%
AA-Omniscience8.416.4
Humanity's Last Exam (AA)49.4%46.5%
SciCode (AA)60.9%58.9%
Terminal-Bench 2.1 (Vals)67.8%
Vals Index59.7
CritPt26.6%20.9%
GDPval (AA)58.7%53.3%
Code Migration43.0%
Excel Modeling Benchmark62.9%
Finance Agent v258.3%
Harvey's Legal Agent Benchmark10.8%
Legal Research Bench47.1%
Terminal-Bench Science2.9%
Vibe Code Bench v1.185.2%
Terminal-Bench 4.0 (AA)34.9%33.3%
AutomationBench58.6%51.0%
GDP.pdf19.2%14.8%
MLCR16.7%
AA-Omniscience: accuracy34.9%41.5%
AA-Omniscience: non-hallucination59.4%57.0%
AA-Briefcase15221432

Data as of 2026-09-23. Best configuration of each model; every score links to its source on the model pages.

MiMo-V2.6-Pro vs Step 5 Preview: questions

Is MiMo-V2.6-Pro better than Step 5 Preview?
MiMo-V2.6-Pro and Step 5 Preview are level on quality (64.3 vs 64.4). The BenchLeader Index combines every independent quality benchmark; Step 5 Preview is ahead overall as of 2026-09-23, but check the category scores for your use.
Which is cheaper, MiMo-V2.6-Pro or Step 5 Preview?
MiMo-V2.6-Pro is cheaper: $0.548 against $1.43 per million tokens, blended at three input tokens per output token.
Which is faster, MiMo-V2.6-Pro or Step 5 Preview?
Step 5 Preview streams faster: 83 against 54 output tokens per second.
Which has the larger context window?
MiMo-V2.6-Pro accepts more context: 1.0M against 1M tokens.