BenchLeader

Claude Sonnet 5 vs Deepseek v4 Pro

Verdict
  • Claude Sonnet 5 (high) and Deepseek v4 Pro (high) are level on quality (60.0 vs 60.4).
  • Claude Sonnet 5 (high) is stronger in agents & tools, coding, composite, knowledge, long context, multimodal, reasoning.
  • Deepseek v4 Pro (high) is stronger in human preference, maths.
MetricClaude Sonnet 5 (high)Deepseek v4 Pro (high)
BenchLeader Index60.060.4
Agents & tools score60.8
Coding score61.559.3
Composite score63.4
Human preference score65.966.0
Knowledge score58.9
Long context score62.1
Multimodal score63.1
Reasoning score66.565.9
Maths score66.2
Blended price $/M$4.00
Output speed68 tok/s
Time to first answer1.8 s
Context window1M
GPQA Diamond90.9%
OTIS Mock AIME95.6%
SciCode48.6%46.4%
WeirdML68.8%46.5%
LMArena Text14611462
LMArena Hard Prompts14881482
LMArena Coding15211505
LMArena WebDev15371581
LMArena Vision1278
LMArena Agent5.9
AA Intelligence Index32.0
AA-LCR76.7%
AA-Omniscience-3.7
Humanity's Last Exam (AA)35.7%
SciCode (AA)54.3%

Data as of 2026-09-13. Best configuration of each model; every score links to its source on the model pages.

Claude Sonnet 5 vs Deepseek v4 Pro: questions

Is Claude Sonnet 5 better than Deepseek v4 Pro?
Claude Sonnet 5 (high) and Deepseek v4 Pro (high) are level on quality (60.0 vs 60.4). The BenchLeader Index combines every independent quality benchmark; Deepseek v4 Pro (high) is ahead overall as of 2026-09-13, but check the category scores for your use.
Is Claude Sonnet 5 better than Deepseek v4 Pro for coding?
Claude Sonnet 5 scores higher in coding (62 vs 59 on the category index, where 50 is average).