BenchLeader

ALE-Bench

Long-horizon algorithm engineering on AtCoder heuristic contests.

As of 19 Sept 2026, GPT-5.6 Sol leads ALE-Bench on BenchLeader with 2176.9, ahead of Claude Opus 5 at 2164.6, across 110 model configurations with a published result.

Published by
Epoch AI Benchmarking Hub
Category
Coding
Index weight
Reference only
Models
110
Data as of
19 Sept 2026

CC BY 4.0 — Epoch AI, ‘AI Benchmarking Hub’, epoch.ai/benchmarks. Mirrored boards credit their original publishers.

What the test looks like

The model iterates on a heuristic-contest solution over hours, scored like a contestant.

How it is scored

Contest performance rating, as published by ALE-Bench.

What to keep in mind

Rating scale like competitive programming; not a percent.

110 of 110
#
1GPT-5.6 SolmaxOpenAI2176.968.82026-07-09
2Claude Opus 5highAnthropic2164.670.22026-07-24
3Claude Fable 5.1highAnthropic2143.272.02026-09-01
4Claude Fable 5highAnthropic2041.364.02026-06-09
5GPT-5.6 TerramaxOpenAI1951.465.02026-07-09
6GPT-5.5xhighOpenAI1943.067.62026-04-23
7GPT-5.6 LunamaxOpenAI1667.460.22026-07-09
8GPT-5.3 CodexxhighOpenAI1655.265.42026-02-05
9GPT-5.4highOpenAI160759.02026-03-05
10GPT-5.5mediumOpenAI1589.464.62026-04-23
11Claude Opus 4.8highAnthropic1563.862.22026-05-28
12Kimi K3maxMoonshot AIopen ↗1524.567.22026-07-16
13GPT-5.4mediumOpenAI1520.755.92026-03-05
14Grok 4.6xhighxAI1508.164.62026-08-12
15Claude Sonnet 5highAnthropic1463.162.22026-06-30
16Claude Opus 4.8no reasoningAnthropic1411.82026-05-28
17DeepSeek V4 Pro 0813maxDeepSeekopen ↗1403.256.02026-08-13
18Gemini 3 FlashGoogle1367.257.72025-12-17
19Claude Sonnet 4.6mediumAnthropic1327.352.02026-02-17
20Claude Opus 4.7Anthropic1323.064.52026-04-16
21GLM-5.3highZhipu AIopen ↗1317.42026-08-14
22Grok 4.5highxAI1309.361.02026-07-08
23DeepSeek V4 Flash 0731maxDeepSeekopen ↗1306.154.42026-07-31
24GPT-5.2-CodexOpenAI1299.954.02025-12-18
25GPT-5.2highOpenAI1293.555.72025-12-11
26GPT-5.2mediumOpenAI1249.858.72025-12-11
27GPT-5.1-CodexOpenAI1244.92025-11-12
28GPT-5.1-Codex-MaxmaxOpenAI1208.82025-11-19
29GPT-5.1highOpenAI1192.258.72025-11-13
30Qwen3 7maxAlibaba1189.461.02026-05-19
31GPT-5.4 minihighOpenAI1188.653.02026-03-17
32Gemini 3 ProGoogle1176.861.12025-11-18
33GPT-5highOpenAI1162.558.62025-08-07
34Gemini 3.1 ProGoogle1160.663.92026-02-19
35Grok 4.20xAI1150.32026-02-17
36GPT-5.5no reasoningOpenAI1127.654.02026-04-23
37Kimi K2.6Moonshot AIopen ↗1092.760.52026-04-20
38GPT-5.4no reasoningOpenAI1086.053.42026-03-05
39GLM-5.2highZhipu AIopen ↗1047.02026-06-16
40Claude Opus 4.5Anthropic1025.458.22025-11-24
41GLM-5.2maxZhipu AIopen ↗1010.263.92026-06-16
42Deepseek v4 ProhighDeepSeek1006.160.82026-04-24
43GPT-5.4 nanohighOpenAI1004.549.12026-03-17
44Claude Opus 4.6Anthropic996.563.72026-02-05
45Grok 4.3xAI944.250.52026-04-17
46o3highOpenAI933.553.02025-04-16
47Gemma 4 26B A4BGoogle927.257.12026-04-02
48Gemma 4 31BGoogleopen ↗925.552.72026-04-02
49Gemini 3.5 FlashhighGoogle911.063.62026-05-19
50Gemini 3.7 FlashhighGoogle904.364.62026-08-13
51MiMo-V2.5-ProXiaomiopen ↗899.859.2
52GLM-5.1Zhipu AIopen ↗887.156.92026-04-07
53Kimi K2.7 CodeMoonshot AIopen ↗886.255.92026-06-12
54o4-minihighOpenAI826.250.82025-04-16
55Kimi K2.5Moonshot AIopen821.653.92026-01-27
56GPT-5minimalOpenAI807.646.02025-08-07
57DeepSeek R1 0528DeepSeekopen804.150.22025-05-28
58GPT-5 minihighOpenAI799.855.32025-08-07
59Gemini 3.1 Flash LiteGoogle797.754.52026-03-03
60Claude Sonnet 4.5Anthropic796.154.32025-09-29

Cite as: BenchLeader, “ALE-Bench leaderboard”, https://www.benchleader.com/benchmarks/ale_bench, data as of 19 Sept 2026.

ALE-Bench: questions

What does ALE-Bench measure?
The model iterates on a heuristic-contest solution over hours, scored like a contestant. Scores are reported in score; higher is better.
Which AI model leads ALE-Bench?
GPT-5.6 Sol leads ALE-Bench with 2176.9 as of 19 Sept 2026, ahead of Claude Opus 5 at 2164.6.
How many models have ALE-Bench results?
110 model configurations have a ALE-Bench result on BenchLeader, all taken from Epoch AI Benchmarking Hub.
Who runs ALE-Bench and how often is it updated?
ALE-Bench is published by Epoch AI Benchmarking Hub. BenchLeader re-reads the published results every morning and records the date each result was published.
Does ALE-Bench count toward the BenchLeader Index?
No. ALE-Bench is shown for reference but left out of the composite index.