BenchLeader
OpenAIReasoning model

o1-mini

Best configuration ranks #388 of 610 on the BenchLeader Index at 45.5 ±6.4. Last measured 2 Sept 2026. Released 12 Sept 2024.

Blended price
Output speed
First answer
Context
128k
How it scores by categoryDashed line = average model (50). One step of 15 = one standard deviation.
  1. Overall index46
  2. Reasoning37
  3. Coding45
  4. Human preference51
  5. Composite41

Reasoning-effort configurations

The same model behaves differently depending on how much it is allowed to think. Each row is one setting, scored only on the benchmarks that were run at that setting. “Default” means the publisher did not say which setting was used.

EffortIndexRankSpeedFirst answerChat reply costCategories
medium44.4#415Agents & tools 38 · Coding 41 · Maths 53 · Reasoning 35
highKnowledge 56 · Maths 55 · Reasoning 47
defaultbest45.5#388Coding 45 · Composite 41 · Human preference 51 · Reasoning 37

Benchmark results

One column per reasoning effort. Rank is among every configuration of every model on that benchmark. Hover a score for the run it came from.

Reasoning

BenchmarkmediumhighdefaultSource
GPQA Diamond59.5%#18062.4%#175Epoch AI Benchmarking Hub
SimpleBench18.1%#82SimpleBench
LMArena Hard Prompts1360#186LMArena
GPQA Diamond (AA)not in index60.3%#355Artificial Analysis
Humanity's Last Exam (AA)not in index3.6%#499Artificial Analysis
ARC-AGI-114.0%#164ARC Prize
ARC-AGI-20.8%#164ARC Prize

Coding

BenchmarkmediumhighdefaultSource
WeirdML36.3%#119WeirdML
LMArena Coding1387#185LMArena
Aider Polyglot32.9%#31Aider polyglot leaderboard

Agents & tools

BenchmarkmediumhighdefaultSource
Cybench10.0%#15Cybench

Maths

BenchmarkmediumhighdefaultSource
OTIS Mock AIME44.7%#17346.9%#168Epoch AI Benchmarking Hub
MATH Level 584.3%#2889.2%#21Epoch AI Benchmarking Hub

Knowledge

BenchmarkmediumhighdefaultSource
MedQA90.2%#46Vals AI

Human preference

BenchmarkmediumhighdefaultSource
LMArena Text1337#193LMArena

Composite

BenchmarkmediumhighdefaultSource
Epoch Capabilities Indexnot in index135.8#115Epoch AI Benchmarking Hub
AA Intelligence Index9.8#320Artificial Analysis

See also

Data as of 9 Sept 2026. Compare these configurations.