BenchLeader

Agnes 3.0 Flash vs MiMo-V2.6-Pro

Verdict
  • MiMo-V2.6-Pro leads on quality: 64.3 vs 61.4.
  • Agnes 3.0 Flash is stronger in agents & tools.
  • MiMo-V2.6-Pro is stronger in composite, knowledge, long context, reasoning.
  • Agnes 3.0 Flash is 7.3× cheaper ($0.075 vs $0.548 per 1M blended).
MetricAgnes 3.0 FlashMiMo-V2.6-Pro
BenchLeader Index61.464.3
Agents & tools score76.855.1
Composite score71.884.8
Knowledge score58.066.8
Long context score66.369.1
Reasoning score68.588.8
Blended price $/M$0.075$0.548
Output speed54 tok/s
Time to first answer39.8 s
Context window1M1.0M
AA Intelligence Index35.546.3
AA-LCR81.0%86.3%
AA-Omniscience-10.68.4
GPQA Diamond (AA)92.4%
Humanity's Last Exam (AA)38.5%49.4%
SciCode (AA)51.6%60.9%
Terminal-Bench 2.1 (Vals)67.8%
Vals Index59.7
CritPt15.1%26.6%
GDPval (AA)52.5%58.7%
τ²-Bench Banking (AA)47.6%
Code Migration43.0%
Excel Modeling Benchmark62.9%
Finance Agent v258.3%
Harvey's Legal Agent Benchmark10.8%
Legal Research Bench47.1%
Terminal-Bench Science2.9%
Vibe Code Bench v1.185.2%
Terminal-Bench 4.0 (AA)7.1%34.9%
Terminal-Bench 2.1 (AA)82.0%
AutomationBench58.6%
GDP.pdf19.2%
AA-Omniscience: accuracy25.5%34.9%
AA-Omniscience: non-hallucination51.6%59.4%
AA-Briefcase1522

Data as of 2026-09-23. Best configuration of each model; every score links to its source on the model pages.

Agnes 3.0 Flash vs MiMo-V2.6-Pro: questions

Is Agnes 3.0 Flash better than MiMo-V2.6-Pro?
MiMo-V2.6-Pro leads on quality: 64.3 vs 61.4. The BenchLeader Index combines every independent quality benchmark; MiMo-V2.6-Pro is ahead overall as of 2026-09-23, but check the category scores for your use.
Is Agnes 3.0 Flash better than MiMo-V2.6-Pro for agentic tasks?
Agnes 3.0 Flash scores higher in agentic tasks (77 vs 55 on the category index, where 50 is average).
Which is cheaper, Agnes 3.0 Flash or MiMo-V2.6-Pro?
Agnes 3.0 Flash is cheaper: $0.075 against $0.548 per million tokens, blended at three input tokens per output token.
Which has the larger context window?
MiMo-V2.6-Pro accepts more context: 1.0M against 1M tokens.