BenchLeader

Dola Seed 2.0 Pro vs GPT-6 Sol

Verdict
  • GPT-6 Sol (max) leads on quality: 68.4 vs 59.2.
  • Dola Seed 2.0 Pro is stronger in human preference, maths.
  • GPT-6 Sol (max) is stronger in coding, multimodal, reasoning, agents & tools, composite, knowledge, long context.
MetricDola Seed 2.0 ProGPT-6 Sol (max)
BenchLeader Index59.268.4
Coding score66.172.3
Human preference score65.2
Maths score64.2
Multimodal score62.367.2
Reasoning score65.595.0
Agents & tools score69.5
Composite score74.6
Knowledge score75.4
Long context score67.8
Blended price $/M$4.00
Output speed116 tok/s
Time to first answer136.1 s
Context window1.1M
LMArena Text1456
LMArena Hard Prompts1480
LMArena Coding1513
LMArena WebDev1686
LMArena Vision1275
LiveBench79.3%
LiveBench Reasoning88.7%
LiveBench Coding81.8%
LiveBench Agentic Coding52.9%
LiveBench Mathematics96.4%
LiveBench Data Analysis81.2%
LiveBench Language85.3%
LiveBench Instruction Following68.6%
AA Intelligence Index47.5
AA-LCR83.7%
MMMU-Pro83.3%
AA-Omniscience27.1
Humanity's Last Exam (AA)47.9%
SciCode (AA)57.6%
IOI82.6%
Terminal-Bench 2.1 (Vals)83.2%
Vals Index62.6
CritPt30.9%
GDPval (AA)49.4%
BioMysteryBench74.8%
Code Migration57.2%
Excel Modeling Benchmark71.5%
Finance Agent v249.0%
Harvey's Legal Agent Benchmark1.7%
Legal Research Bench28.9%
MedCode47.1%
MedScribe82.0%
ProgramBench2.0%
Public Benefits Bench56.6%
SAGE44.8%
Tax Agent Bench53.0%
Vibe Code Bench v1.187.8%
LMArena Maths1451
LMArena Creative Writing1402
LMArena Instruction Following1434
LMArena Multi-turn1448
LMArena Longer Queries1451
EBR-bench53.3%
Terminal-Bench 4.0 (AA)43.9%
AutomationBench61.6%
GDP.pdf24.8%
MLCR16.1%
AA-Omniscience: accuracy54.5%
AA-Omniscience: non-hallucination39.9%
AA-Briefcase1483

Data as of 2026-09-24. Best configuration of each model; every score links to its source on the model pages.

Dola Seed 2.0 Pro vs GPT-6 Sol: questions

Is Dola Seed 2.0 Pro better than GPT-6 Sol?
GPT-6 Sol (max) leads on quality: 68.4 vs 59.2. The BenchLeader Index combines every independent quality benchmark; GPT-6 Sol (max) is ahead overall as of 2026-09-24, but check the category scores for your use.
Is Dola Seed 2.0 Pro better than GPT-6 Sol for coding?
GPT-6 Sol scores higher in coding (72 vs 66 on the category index, where 50 is average).