BenchLeader

Grok 4.3 vs Mimo v2 Pro

Verdict
  • Grok 4.3 (medium) and Mimo v2 Pro are level on quality (60.7 vs 60.6).
  • Grok 4.3 (medium) is stronger in instruction following, knowledge, long context, multimodal.
  • Mimo v2 Pro is stronger in agents & tools, composite, coding, human preference, reasoning.
  • Grok 4.3 (medium) is 1.9× cheaper ($1.56 vs $3.00 per 1M blended).
MetricGrok 4.3 (medium)Mimo v2 Pro
BenchLeader Index60.760.6
Agents & tools score60.169.4
Composite score60.365.2
Instruction following score80.067.5
Knowledge score72.166.3
Long context score64.060.5
Multimodal score60.2
Coding score54.7
Human preference score64.5
Reasoning score65.2
Blended price $/M$1.56$3.00
Output speed112 tok/s
Time to first answer12.1 s
Context window1M1.0M
LMArena Text1448
LMArena Hard Prompts1476
LMArena Coding1503
LMArena WebDev1433
AA Intelligence Index24.828.6
IFBench83.3%68.8%
AA-LCR75.0%68.3%
MMMU-Pro75.8%
AA-Omniscience16.74.6
Terminal-Bench Hard30.3%40.9%
GPQA Diamond (AA)89.0%87.0%
Humanity's Last Exam (AA)30.0%30.4%
τ²-Bench Telecom (AA)91.2%95.0%

Data as of 2026-09-10. Best configuration of each model; every score links to its source on the model pages.

Grok 4.3 vs Mimo v2 Pro: questions

Is Grok 4.3 better than Mimo v2 Pro?
Grok 4.3 (medium) and Mimo v2 Pro are level on quality (60.7 vs 60.6). The BenchLeader Index combines every independent quality benchmark; Grok 4.3 (medium) is ahead overall as of 2026-09-10, but check the category scores for your use.
Is Grok 4.3 better than Mimo v2 Pro for agentic tasks?
Mimo v2 Pro scores higher in agentic tasks (69 vs 60 on the category index, where 50 is average).
Which is cheaper, Grok 4.3 or Mimo v2 Pro?
Grok 4.3 is cheaper: $1.56 against $3.00 per million tokens, blended at three input tokens per output token.
Which has the larger context window?
Mimo v2 Pro accepts more context: 1.0M against 1M tokens.