BenchLeader

Claude Haiku 5.5 vs Claude Opus 5

Verdict
  • Claude Opus 5 (high) leads on quality: 68.3 vs 61.3.
  • Claude Opus 5 (high) is stronger in coding, composite, knowledge, long context, maths, reasoning, agents & tools, human preference, multimodal.
  • Claude Haiku 5.5 (high) is 50× cheaper ($0.200 vs $10.00 per 1M blended).
  • Claude Haiku 5.5 (high) streams 3.2× faster (174 vs 54 tokens per second).
MetricClaude Haiku 5.5 (high)Claude Opus 5 (high)
BenchLeader Index61.368.3
Coding score62.669.2
Composite score72.484.1
Knowledge score64.376.9
Long context score63.864.7
Maths score66.271.1
Reasoning score72.172.7
Agents & tools score–65.1
Human preference score–68.9
Multimodal score–66.4
Blended price $/M$0.200$10.00
Output speed174 tok/s54 tok/s
Time to first answer22.6 s20.4 s
Context window1M1M
OTIS Mock AIME97.2%–
Terminal-Bench–50.3%
OSWorld-Verified 2.0–29.0%
SciCode–54.3%
WeirdML–91.6%
LMArena Text–1490
LMArena Hard Prompts–1515
LMArena Coding–1534
LMArena WebDev15871657
LMArena Vision–1319
LMArena Agent–8
AA Intelligence Index v4.3.237.848.1
AA-LCR77.3%79.0%
MMMU-Pro–82.4%
AA-Omniscience5.833.7
GPQA Diamond (AA)–93.7%
Humanity's Last Exam (AA)37.3%52.8%
SciCode (AA)48.7%55.4%
ARC-AGI-1–97.5%
ARC-AGI-2–88.3%
ARC-AGI-3–30.2%
CritPt18.6%28.3%
GDPval-AA v2.145.9%54.7%
τ³-Banking (AA)–44.7%
LMArena Maths–1521
LMArena Creative Writing–1471
LMArena Instruction Following–1497
LMArena Multi-turn–1484
LMArena Longer Queries–1504
LMArena Document–1490
DeepSWE v1.1–72.8%
LMCA–63.7%
DTBench–97.9%
CursorBench–44.7%
ALE-Bench–2164.6
Terminal-Bench 4.0 (AA)21.7%46.0%
Terminal-Bench 2.1 (AA)–87.6%
AutomationBench33.7%53.6%
GDP.pdf17.2%19.6%
MLCR–59.4%
Harvey LAB1.1%–
AA-Omniscience: accuracy34.8%58.9%
AA-Omniscience: non-hallucination55.5%38.8%
AA-Briefcase v1.114421561

Data as of 2026-10-11. Best configuration of each model; every score links to its source on the model pages.

Claude Haiku 5.5 vs Claude Opus 5: questions

Is Claude Haiku 5.5 better than Claude Opus 5?
Claude Opus 5 (high) leads on quality: 68.3 vs 61.3. The BenchLeader Index combines every independent quality benchmark; Claude Opus 5 (high) is ahead overall as of 2026-10-11, but check the category scores for your use.
Is Claude Haiku 5.5 better than Claude Opus 5 for coding?
Claude Opus 5 scores higher in coding (69 vs 63 on the category index, where 50 is average).
Which is cheaper, Claude Haiku 5.5 or Claude Opus 5?
Claude Haiku 5.5 is cheaper: $0.200 against $10.00 per million tokens, blended at three input tokens per output token.
Which is faster, Claude Haiku 5.5 or Claude Opus 5?
Claude Haiku 5.5 streams faster: 174 against 54 output tokens per second.
Which has the larger context window?
Both accept 1M tokens of context.