BenchLeader

Deepseek v4 Pro High vs GPT-6 Sol

Verdict
  • GPT-6 Sol (max) leads on quality: 67.7 vs 60.8.
  • Deepseek v4 Pro High (high) is stronger in coding, human preference, maths.
  • GPT-6 Sol (max) is stronger in reasoning, composite, knowledge, long context, multimodal.
MetricDeepseek v4 Pro High (high)GPT-6 Sol (max)
BenchLeader Index60.867.7
Coding score59.0
Human preference score66.2
Maths score66.2
Reasoning score66.095.0
Composite score86.2
Knowledge score75.4
Long context score67.7
Multimodal score67.1
Blended price $/M$4.00
Output speed126 tok/s
Time to first answer107.2 s
Context window1.1M
GPQA Diamond90.9%
OTIS Mock AIME95.6%
SciCode46.4%
WeirdML46.5%
LMArena Text1463
LMArena Hard Prompts1482
LMArena Coding1503
LMArena WebDev1581
AA Intelligence Index47.5
AA-LCR83.7%
MMMU-Pro83.3%
AA-Omniscience27.1
Humanity's Last Exam (AA)47.9%
SciCode (AA)57.6%
CritPt30.9%
GDPval (AA)49.4%
LMArena Maths1469
LMArena Creative Writing1442
LMArena Instruction Following1457
LMArena Multi-turn1469
LMArena Longer Queries1474
Chess Puzzles13.0%
CL-bench Life13.5%
Surface Evolver Bench40.0%
ALE-Bench1006.1
Terminal-Bench 4.0 (AA)43.9%
AutomationBench61.6%
GDP.pdf24.8%
MLCR16.1%
AA-Omniscience: accuracy54.5%
AA-Omniscience: non-hallucination39.9%
AA-Briefcase1483

Data as of 2026-09-23. Best configuration of each model; every score links to its source on the model pages.

Deepseek v4 Pro High vs GPT-6 Sol: questions

Is Deepseek v4 Pro High better than GPT-6 Sol?
GPT-6 Sol (max) leads on quality: 67.7 vs 60.8. The BenchLeader Index combines every independent quality benchmark; GPT-6 Sol (max) is ahead overall as of 2026-09-23, but check the category scores for your use.