CL-bench Life
The everyday-life subset of CL-bench.
As of 19 Sept 2026, GPT-5.5 leads CL-bench Life on BenchLeader with 22.2%, ahead of GPT-5.4 at 21.7%, across 16 model configurations with a published result.
- Published by
- Epoch AI Benchmarking Hub
- Category
- Reasoning
- Index weight
- Reference only
- Models
- 16
- Data as of
- 19 Sept 2026
CC BY 4.0 — Epoch AI, ‘AI Benchmarking Hub’, epoch.ai/benchmarks. Mirrored boards credit their original publishers.
What the test looks like
CL-bench tasks drawn from everyday situations.
How it is scored
Overall score, as published.
What to keep in mind
New; small set.
- 1GPT-5.5 (high)22.2%
- 2GPT-5.4 (xhigh)21.7%
- 3GPT-5.4 (high)19.3%
- 4GPT-5.1 (high)17.3%
- 5Claude Opus 4.6 (high)17.0%
- 6Gemini 3.1 Pro16.9%
- 7GPT-5.413.8%
- 8Claude Opus 4.613.6%
- 9Deepseek v4 Pro (high)13.5%
- 10Kimi K2.513.2%
- 11Qwen3.5 Plus12.4%
- 12Grok 4.2011.9%
- 13GLM-4.710.9%
- 14DeepSeek V3.2 (thinking)9.5%
- 15Mimo v2 Pro6.9%
16 of 16
| # | ||||
|---|---|---|---|---|
| 1 | 22.2% | 67.0 | 2026-04-23 | |
| 2 | 21.7% | 65.4 | 2026-03-05 | |
| 3 | 19.3% | 59.0 | 2026-03-05 | |
| 4 | 17.3% | 58.7 | 2025-11-13 | |
| 5 | 17.0% | 61.3 | 2026-02-05 | |
| 6 | 16.9% | 63.9 | 2026-02-19 | |
| 7 | 13.8% | 59.2 | 2026-03-05 | |
| 8 | 13.6% | 63.7 | 2026-02-05 | |
| 9 | 13.5% | 60.8 | 2026-04-24 | |
| 10 | 13.2% | 53.9 | 2026-01-27 | |
| 11 | 12.4% | – | 2026-02-16 | |
| 12 | 11.9% | – | 2026-02-17 | |
| 13 | 10.9% | 52.4 | 2025-12-22 | |
| 14 | 9.5% | 56.1 | 2025-09-29 | |
| 15 | 6.9% | 59.8 | – | |
| 16 | 6.3% | 54.2 | 2026-02-12 |
Cite as: BenchLeader, “CL-bench Life leaderboard”, https://www.benchleader.com/benchmarks/cl_bench_life, data as of 19 Sept 2026.
CL-bench Life: questions
- What does CL-bench Life measure?
- CL-bench tasks drawn from everyday situations. Scores are reported in percent of tasks solved; higher is better.
- Which AI model leads CL-bench Life?
- GPT-5.5 leads CL-bench Life with 22.2% as of 19 Sept 2026, ahead of GPT-5.4 at 21.7%.
- How many models have CL-bench Life results?
- 16 model configurations have a CL-bench Life result on BenchLeader, all taken from Epoch AI Benchmarking Hub.
- Who runs CL-bench Life and how often is it updated?
- CL-bench Life is published by Epoch AI Benchmarking Hub. BenchLeader re-reads the published results every morning and records the date each result was published.
- Does CL-bench Life count toward the BenchLeader Index?
- No. CL-bench Life is shown for reference but left out of the composite index.