BenchLeader

Claude Sonnet 5 vs GPT-6 Astra

Verdict
  • GPT-6 Astra (max) leads on quality: 70.9 vs 61.4.
  • Claude Sonnet 5 (high) is stronger in human preference, long context, multimodal.
  • GPT-6 Astra (max) is stronger in agents & tools, coding, composite, knowledge, reasoning, maths.
  • Claude Sonnet 5 (high) is 5.0× cheaper ($4.00 vs $20.00 per 1M blended).
  • Claude Sonnet 5 (high) streams 2.2× faster (60 vs 27 tokens per second).
MetricClaude Sonnet 5 (high)GPT-6 Astra (max)
BenchLeader Index61.470.9
Agents & tools score61.868.2
Coding score61.678.6
Composite score69.572.9
Human preference score66.2
Knowledge score62.480.0
Long context score64.8
Multimodal score63.1
Reasoning score66.873.6
Maths score77.2
Blended price $/M$4.00$20.00
Output speed60 tok/s27 tok/s
Time to first answer10.0 s337.3 s
Context window1M1.1M
GPQA Diamond95.8%
FrontierMath Tiers 1–393.7%
FrontierMath Tier 497.6%
OTIS Mock AIME100.0%
SimpleQA Verified75.6%
Terminal-Bench58.2%
SciCode48.6%56.5%
WeirdML68.8%
FrontierCode53.3%
LMArena Text1462
LMArena Hard Prompts1490
LMArena Coding1521
LMArena WebDev15381796
LMArena Vision1278
LMArena Agent6.312.5
LiveBench82.2%
LiveBench Reasoning92.7%
LiveBench Coding80.4%
LiveBench Agentic Coding57.3%
LiveBench Mathematics96.8%
LiveBench Data Analysis83.0%
LiveBench Language89.4%
AA Intelligence Index32.0
AA-LCR76.7%
AA-Omniscience-3.7
Humanity's Last Exam (AA)35.7%
SciCode (AA)54.3%
Terminal-Bench 2.1 (Vals)87.3%
Vals Index66.6
ARC-AGI-197.5%
ARC-AGI-295.0%
ARC-AGI-398.6%

Data as of 2026-09-10. Best configuration of each model; every score links to its source on the model pages.

Claude Sonnet 5 vs GPT-6 Astra: questions

Is Claude Sonnet 5 better than GPT-6 Astra?
GPT-6 Astra (max) leads on quality: 70.9 vs 61.4. The BenchLeader Index combines every independent quality benchmark; GPT-6 Astra (max) is ahead overall as of 2026-09-10, but check the category scores for your use.
Is Claude Sonnet 5 better than GPT-6 Astra for coding?
GPT-6 Astra scores higher in coding (79 vs 62 on the category index, where 50 is average).
Is Claude Sonnet 5 better than GPT-6 Astra for agentic tasks?
GPT-6 Astra scores higher in agentic tasks (68 vs 62 on the category index, where 50 is average).
Which is cheaper, Claude Sonnet 5 or GPT-6 Astra?
Claude Sonnet 5 is cheaper: $4.00 against $20.00 per million tokens, blended at three input tokens per output token.
Which is faster, Claude Sonnet 5 or GPT-6 Astra?
Claude Sonnet 5 streams faster: 60 against 27 output tokens per second.
Which has the larger context window?
GPT-6 Astra accepts more context: 1.1M against 1M tokens.