ALE-Bench
Long-horizon algorithm engineering on AtCoder heuristic contests.
As of 19 Sept 2026, GPT-5.6 Sol leads ALE-Bench on BenchLeader with 2176.9, ahead of Claude Opus 5 at 2164.6, across 110 model configurations with a published result.
- Published by
- Epoch AI Benchmarking Hub
- Category
- Coding
- Index weight
- Reference only
- Models
- 110
- Data as of
- 19 Sept 2026
CC BY 4.0 — Epoch AI, ‘AI Benchmarking Hub’, epoch.ai/benchmarks. Mirrored boards credit their original publishers.
What the test looks like
The model iterates on a heuristic-contest solution over hours, scored like a contestant.
How it is scored
Contest performance rating, as published by ALE-Bench.
What to keep in mind
Rating scale like competitive programming; not a percent.
- 1GPT-5.6 Sol (max)2176.9
- 2Claude Opus 5 (high)2164.6
- 3Claude Fable 5.1 (high)2143.2
- 4Claude Fable 5 (high)2041.3
- 5GPT-5.6 Terra (max)1951.4
- 6GPT-5.5 (xhigh)1943.0
- 7GPT-5.6 Luna (max)1667.4
- 8GPT-5.3 Codex (xhigh)1655.2
- 9GPT-5.4 (high)1607
- 10GPT-5.5 (medium)1589.4
- 11Claude Opus 4.8 (high)1563.8
- 12Kimi K3 (max)1524.5
- 13GPT-5.4 (medium)1520.7
- 14Grok 4.6 (xhigh)1508.1
- 15Claude Sonnet 5 (high)1463.1
110 of 110
| # | ||||
|---|---|---|---|---|
| 1 | 2176.9 | 68.8 | 2026-07-09 | |
| 2 | 2164.6 | 70.2 | 2026-07-24 | |
| 3 | 2143.2 | 72.0 | 2026-09-01 | |
| 4 | 2041.3 | 64.0 | 2026-06-09 | |
| 5 | 1951.4 | 65.0 | 2026-07-09 | |
| 6 | 1943.0 | 67.6 | 2026-04-23 | |
| 7 | 1667.4 | 60.2 | 2026-07-09 | |
| 8 | 1655.2 | 65.4 | 2026-02-05 | |
| 9 | 1607 | 59.0 | 2026-03-05 | |
| 10 | 1589.4 | 64.6 | 2026-04-23 | |
| 11 | 1563.8 | 62.2 | 2026-05-28 | |
| 12 | 1524.5 | 67.2 | 2026-07-16 | |
| 13 | 1520.7 | 55.9 | 2026-03-05 | |
| 14 | 1508.1 | 64.6 | 2026-08-12 | |
| 15 | 1463.1 | 62.2 | 2026-06-30 | |
| 16 | 1411.8 | – | 2026-05-28 | |
| 17 | 1403.2 | 56.0 | 2026-08-13 | |
| 18 | 1367.2 | 57.7 | 2025-12-17 | |
| 19 | 1327.3 | 52.0 | 2026-02-17 | |
| 20 | 1323.0 | 64.5 | 2026-04-16 | |
| 21 | 1317.4 | – | 2026-08-14 | |
| 22 | 1309.3 | 61.0 | 2026-07-08 | |
| 23 | 1306.1 | 54.4 | 2026-07-31 | |
| 24 | 1299.9 | 54.0 | 2025-12-18 | |
| 25 | 1293.5 | 55.7 | 2025-12-11 | |
| 26 | 1249.8 | 58.7 | 2025-12-11 | |
| 27 | 1244.9 | – | 2025-11-12 | |
| 28 | 1208.8 | – | 2025-11-19 | |
| 29 | 1192.2 | 58.7 | 2025-11-13 | |
| 30 | 1189.4 | 61.0 | 2026-05-19 | |
| 31 | 1188.6 | 53.0 | 2026-03-17 | |
| 32 | 1176.8 | 61.1 | 2025-11-18 | |
| 33 | 1162.5 | 58.6 | 2025-08-07 | |
| 34 | 1160.6 | 63.9 | 2026-02-19 | |
| 35 | 1150.3 | – | 2026-02-17 | |
| 36 | 1127.6 | 54.0 | 2026-04-23 | |
| 37 | 1092.7 | 60.5 | 2026-04-20 | |
| 38 | 1086.0 | 53.4 | 2026-03-05 | |
| 39 | 1047.0 | – | 2026-06-16 | |
| 40 | 1025.4 | 58.2 | 2025-11-24 | |
| 41 | 1010.2 | 63.9 | 2026-06-16 | |
| 42 | 1006.1 | 60.8 | 2026-04-24 | |
| 43 | 1004.5 | 49.1 | 2026-03-17 | |
| 44 | 996.5 | 63.7 | 2026-02-05 | |
| 45 | 944.2 | 50.5 | 2026-04-17 | |
| 46 | 933.5 | 53.0 | 2025-04-16 | |
| 47 | 927.2 | 57.1 | 2026-04-02 | |
| 48 | 925.5 | 52.7 | 2026-04-02 | |
| 49 | 911.0 | 63.6 | 2026-05-19 | |
| 50 | 904.3 | 64.6 | 2026-08-13 | |
| 51 | 899.8 | 59.2 | – | |
| 52 | 887.1 | 56.9 | 2026-04-07 | |
| 53 | 886.2 | 55.9 | 2026-06-12 | |
| 54 | 826.2 | 50.8 | 2025-04-16 | |
| 55 | 821.6 | 53.9 | 2026-01-27 | |
| 56 | 807.6 | 46.0 | 2025-08-07 | |
| 57 | 804.1 | 50.2 | 2025-05-28 | |
| 58 | 799.8 | 55.3 | 2025-08-07 | |
| 59 | 797.7 | 54.5 | 2026-03-03 | |
| 60 | 796.1 | 54.3 | 2025-09-29 |
Cite as: BenchLeader, “ALE-Bench leaderboard”, https://www.benchleader.com/benchmarks/ale_bench, data as of 19 Sept 2026.
ALE-Bench: questions
- What does ALE-Bench measure?
- The model iterates on a heuristic-contest solution over hours, scored like a contestant. Scores are reported in score; higher is better.
- Which AI model leads ALE-Bench?
- GPT-5.6 Sol leads ALE-Bench with 2176.9 as of 19 Sept 2026, ahead of Claude Opus 5 at 2164.6.
- How many models have ALE-Bench results?
- 110 model configurations have a ALE-Bench result on BenchLeader, all taken from Epoch AI Benchmarking Hub.
- Who runs ALE-Bench and how often is it updated?
- ALE-Bench is published by Epoch AI Benchmarking Hub. BenchLeader re-reads the published results every morning and records the date each result was published.
- Does ALE-Bench count toward the BenchLeader Index?
- No. ALE-Bench is shown for reference but left out of the composite index.