BenchLeader

Claude 3.5 Haiku

Best configuration ranks #518 of 610 on the BenchLeader Index at 39.8 ±3.7. Last measured 2 Sept 2026. Released 4 Nov 2024.

Blended price
Output speed
First answer
Context
200k
How it scores by categoryDashed line = average model (50). One step of 15 = one standard deviation.
  1. Overall index40
  2. Reasoning37
  3. Coding35
  4. Agents & tools36
  5. Maths33
  6. Knowledge37
  7. Instruction following45
  8. Human preference49
  9. Multimodal34
  10. Long context39
  11. Composite40

Benchmark results

One column per reasoning effort. Rank is among every configuration of every model on that benchmark. Hover a score for the run it came from.

Agents & tools

Knowledge

BenchmarkdefaultSource
AA-Omniscience-22.5#192Artificial Analysis
MMLU-Pro64.1%#123Vals AI
LegalBench70.3%#113Vals AI
CorpFin50.8%#101Vals AI
TaxEval57.4%#130Vals AI

Instruction following

BenchmarkdefaultSource
IFBench42.8%#229Artificial Analysis

Human preference

BenchmarkdefaultSource
LMArena Text1324#207LMArena

Multimodal

BenchmarkdefaultSource
LMArena Vision1092#108LMArena
MMMU-Pro45.6%#218Artificial Analysis

Long context

BenchmarkdefaultSource
AA-LCR27.3%#333Artificial Analysis

See also

Data as of 9 Sept 2026. Compare with another model.