BenchLeader

Grok 4.7 vs MiMo-V2.6-Pro

Verdict
  • MiMo-V2.6-Pro leads on quality: 64.3 vs 62.5.
  • Grok 4.7 (xhigh) is stronger in coding, knowledge.
  • MiMo-V2.6-Pro is stronger in agents & tools, composite, long context, reasoning.
  • MiMo-V2.6-Pro is 5.5× cheaper ($0.548 vs $3.00 per 1M blended).
  • MiMo-V2.6-Pro streams 1.4× faster (54 vs 39 tokens per second).
MetricGrok 4.7 (xhigh)MiMo-V2.6-Pro
BenchLeader Index62.564.3
Agents & tools score50.555.1
Coding score64.7
Composite score71.284.8
Knowledge score71.666.8
Long context score64.169.1
Reasoning score73.088.8
Blended price $/M$3.00$0.548
Output speed39 tok/s54 tok/s
Time to first answer0.9 s39.8 s
Context window500k1.0M
Terminal-Bench37.6%
SciCode57.4%
LMArena WebDev1633
LiveBench77.4%
LiveBench Reasoning82.7%
LiveBench Coding77.2%
LiveBench Agentic Coding54.0%
LiveBench Mathematics95.7%
LiveBench Data Analysis76.9%
LiveBench Language80.1%
LiveBench Instruction Following75.3%
AA Intelligence Index46.546.3
AA-LCR76.7%86.3%
AA-Omniscience32.08.4
Humanity's Last Exam (AA)43.1%49.4%
SciCode (AA)57.4%60.9%
IOI57.7%
LegalBench84.4%
Terminal-Bench 2.1 (Vals)73.4%67.8%
Vals Index60.259.7
CritPt17.7%26.6%
GDPval (AA)59.8%58.7%
BioMysteryBench69.3%
Code Migration44.8%43.0%
Excel Modeling Benchmark67.0%62.9%
Finance Agent v252.3%58.3%
Harvey's Legal Agent Benchmark12.5%10.8%
Legal Research Bench47.1%47.1%
MedCode49.5%
MedScribe89.4%
MysteryMechanism25.2%
ProgramBench0.5%
Public Benefits Bench65.6%
SAGE40.8%
Tax Agent Bench65.6%
Terminal-Bench 4.0 (Vals)28.3%
Terminal-Bench Science11.4%2.9%
Vibe Code Bench v1.186.2%85.2%
CursorBench46.3%
Terminal-Bench 4.0 (AA)25.8%34.9%
AutomationBench65.6%58.6%
GDP.pdf20.0%19.2%
MLCR15.0%
AA-Omniscience: accuracy47.5%34.9%
AA-Omniscience: non-hallucination70.7%59.4%
AA-Briefcase16571522

Data as of 2026-09-23. Best configuration of each model; every score links to its source on the model pages.

Grok 4.7 vs MiMo-V2.6-Pro: questions

Is Grok 4.7 better than MiMo-V2.6-Pro?
MiMo-V2.6-Pro leads on quality: 64.3 vs 62.5. The BenchLeader Index combines every independent quality benchmark; MiMo-V2.6-Pro is ahead overall as of 2026-09-23, but check the category scores for your use.
Is Grok 4.7 better than MiMo-V2.6-Pro for agentic tasks?
MiMo-V2.6-Pro scores higher in agentic tasks (55 vs 51 on the category index, where 50 is average).
Which is cheaper, Grok 4.7 or MiMo-V2.6-Pro?
MiMo-V2.6-Pro is cheaper: $0.548 against $3.00 per million tokens, blended at three input tokens per output token.
Which is faster, Grok 4.7 or MiMo-V2.6-Pro?
MiMo-V2.6-Pro streams faster: 54 against 39 output tokens per second.
Which has the larger context window?
MiMo-V2.6-Pro accepts more context: 1.0M against 500k tokens.