BenchLeader

Find a model

Most choices start from constraints rather than a ranking: a budget, a wait you can live with, a licence you need, one thing the model has to be good at. Set those and the shortlist ranks on the skill you asked for, not on the overall index.

434 ranked models · data as of 11 Oct 2026

Good at
Costs at most
Runs at least
Answers within
Context of
And

434 models match, highest index is Claude Opus 5.5 at 72.0 for $8.00 per million tokens.

#ModelIndexPrice $/MSpeedFirst answerContext
1Claude Opus 5.5maxAnthropic72.0±4$8.0096662 s1M
2Claude Fable 5.1maxAnthropic70.8±3$20.0070291 s1M
3GPT-6 AstramaxOpenAI70.2±3$20.0047384 s1.1M
4Gemini 4 ArgonhighGoogle70.0±4$4.00––1M
5GPT-6.1 SolmaxOpenAI69.3±3$4.0056327 s1.1M
6Claude Fable 5maxAnthropic69.1±4$20.0065118 s1M
7Claude Opus 5highAnthropic68.3±4$10.005420 s1M
8Muse Spark 1.3maxMeta67.7±4$2.0017571 s1.0M
9Claude Sonnet 5.5maxAnthropic67.6±6$4.00141487 s1M
10GPT-5.6 SolmaxOpenAI67.5±4$8.007485 s1.1M
11GPT-5.5xhighOpenAI65.9±4$11.258839 s1.1M
12GPT-6 SolmaxOpenAI65.6±3$4.0086130 s1.1M
13Kimi K3maxMoonshot AIopen65.5±3$6.004152 s1.0M
14GPT-5.5 ProxhighOpenAI65.5±9$67.50223.39 s1.1M
15MiMo-V2.6-ProXiaomiopen64.8±5$0.5484351 s1.0M
16Muse SparkMeta64.6±3–––262k
17Gemini 3.8 FlashmediumGoogle64.6±5$1.50762.50 s1.0M
18Grok 4.7highSpaceXAI64.3±7$3.006450 s500k
19GPT-5.4xhighOpenAI64.2±3$5.6391155 s1.1M
20Qwen3.8 MaxmaxAlibaba64.2±3$3.003759 s1M
21Grok 4.6mediumSpaceXAI64.2±5$3.005534 s500k
22GLM 5.3maxZhipu AIopen64.0±3$2.158327 s1M
23GPT-5.3-CodexxhighOpenAI63.9±5$4.819170 s400k
24GPT-5.6 TerramaxOpenAI63.7±5$4.50108114 s1.1M
25Claude Opus 4.7Anthropic63.7±2$10.00811.54 s1M
26GPT-5.4 ProOpenAI63.6±3$67.5066.96 s1.1M
27Gemini 3.7 FlashmediumGoogle63.3±3$1.502856.62 s1.0M
28Qwen3.8 2.4T A95BAlibabaopen63.3±5$3.003758 s1M
29Claude Opus 4.6Anthropic63.1±5$10.00401.63 s1M
30Muse Spark 1.1Meta63.1±6$2.002011.92 s1.0M
31Gemini 3.5 FlashmediumGoogle63.0±3$3.3820614 s1.0M
32DeepSeek V4 PromaxDeepSeekopen62.8±3$1.989349 s1M
33Muse Spark 1.2xhighMeta62.7±4$2.0032518 s1.0M
34GLM 5.2maxZhipu AIopen62.6±4$2.158228 s1M
35Qwen3.6 MaxmaxAlibaba62.5±3$2.927131 s262k
36Claude Opus 4.8maxAnthropic62.4±5$10.005834 s1M
37Step 5 PreviewStepFun62.4±5$1.438726 s1M
38GLM 5.3 FlashZhipu AIopen62.2±5$0.2385341 s1M
39Gemini 3.1 ProGoogle62.1±4$4.5011425 s1.0M
40DeepSeek V4.1 FlashmaxDeepSeekopen61.8±3$0.52521710 s1M

Category scores are on the same scale as the index: 50 is the average evaluated model and 15 points is one standard deviation. A model only appears under a skill if it has been measured on benchmarks in that category.