BenchLeader

GPT-5.4 Pro vs MiMo-V2.6-Pro

Verdict
  • GPT-5.4 Pro and MiMo-V2.6-Pro are level on quality (64.0 vs 64.3).
  • GPT-5.4 Pro is stronger in instruction following, knowledge, multimodal.
  • MiMo-V2.6-Pro is stronger in reasoning, agents & tools, composite, long context.
  • MiMo-V2.6-Pro is 123× cheaper ($0.548 vs $67.50 per 1M blended).
  • MiMo-V2.6-Pro streams 27.0× faster (54 vs 2 tokens per second).
MetricGPT-5.4 ProMiMo-V2.6-Pro
BenchLeader Index64.064.3
Instruction following score67.3
Knowledge score73.666.8
Multimodal score69.8
Reasoning score73.288.8
Agents & tools score55.1
Composite score84.8
Long context score69.1
Blended price $/M$67.50$0.548
Output speed2 tok/s54 tok/s
Time to first answer5.4 s39.8 s
Context window1.1M1.0M
Humanity's Last Exam44.3%
SimpleBench74.1%
Epoch Capabilities Index159.1
AA Intelligence Index46.3
AA-LCR86.3%
AA-Omniscience8.4
Humanity's Last Exam (AA)49.4%
SciCode (AA)60.9%
Terminal-Bench 2.1 (Vals)67.8%
Vals Index59.7
MultiChallenge69.2%
VISTA53.9%
MultiNRC62.3%
TutorBench56.6%
CritPt26.6%
GDPval (AA)58.7%
Code Migration43.0%
Excel Modeling Benchmark62.9%
Finance Agent v258.3%
Harvey's Legal Agent Benchmark10.8%
Legal Research Bench47.1%
Terminal-Bench Science2.9%
Vibe Code Bench v1.185.2%
FORTRESS14.8%
MASK91.7%
Terminal-Bench 4.0 (AA)34.9%
AutomationBench58.6%
GDP.pdf19.2%
AA-Omniscience: accuracy34.9%
AA-Omniscience: non-hallucination59.4%
AA-Briefcase1522

Data as of 2026-09-23. Best configuration of each model; every score links to its source on the model pages.

GPT-5.4 Pro vs MiMo-V2.6-Pro: questions

Is GPT-5.4 Pro better than MiMo-V2.6-Pro?
GPT-5.4 Pro and MiMo-V2.6-Pro are level on quality (64.0 vs 64.3). The BenchLeader Index combines every independent quality benchmark; MiMo-V2.6-Pro is ahead overall as of 2026-09-23, but check the category scores for your use.
Which is cheaper, GPT-5.4 Pro or MiMo-V2.6-Pro?
MiMo-V2.6-Pro is cheaper: $0.548 against $67.50 per million tokens, blended at three input tokens per output token.
Which is faster, GPT-5.4 Pro or MiMo-V2.6-Pro?
MiMo-V2.6-Pro streams faster: 54 against 2 output tokens per second.
Which has the larger context window?
GPT-5.4 Pro accepts more context: 1.1M against 1.0M tokens.