BenchLeader

Claude Fable 5.1 vs GLM-5

Verdict
  • Claude Fable 5.1 (max) leads on quality: 70.3 vs 59.1.
  • Claude Fable 5.1 (max) is stronger in agents & tools, coding, composite, human preference, knowledge, maths, reasoning.
  • GLM-5 is stronger in instruction following, long context.
  • GLM-5 is 13× cheaper ($1.55 vs $20.00 per 1M blended).
  • GLM-5 streams 1.6× faster (66 vs 40 tokens per second).
MetricClaude Fable 5.1 (max)GLM-5
BenchLeader Index70.359.1
Agents & tools score67.456.8
Coding score78.452.9
Composite score76.664.3
Human preference score71.365.6
Knowledge score75.964.3
Maths score77.053.1
Reasoning score71.951.5
Instruction following score70.5
Long context score64.3
Blended price $/M$20.00$1.55
Output speed40 tok/s66 tok/s
Time to first answer299.8 s48.7 s
Context window1M205k
GPQA Diamond87.8%
FrontierMath Tiers 1–390.2%
FrontierMath Tier 487.8%
OTIS Mock AIME100.0%80.0%
SWE-bench Verified (Epoch)72.1%
SimpleQA Verified70.8%
Terminal-Bench57.9%52.4%
SimpleBench53.2%
SciCode62.0%
WeirdML92.9%48.2%
APEX-Agents17.2%
ProofBench100.0%
Epoch Capabilities Index145.9
LMArena Text15041458
LMArena Hard Prompts15221478
LMArena Coding15171498
LMArena WebDev17641436
LMArena Agent14.5
LiveBench83.4%
LiveBench Reasoning91.7%
LiveBench Coding86.4%
LiveBench Agentic Coding66.1%
LiveBench Mathematics97.0%
LiveBench Data Analysis80.3%
LiveBench Language89.5%
AA Intelligence Index27.9
IFBench72.3%
AA-LCR75.7%
AA-Omniscience0.3
Terminal-Bench Hard43.2%
GPQA Diamond (AA)82.0%
Humanity's Last Exam (AA)29.3%
τ²-Bench Telecom (AA)98.3%
AIME 202696.7%
HMMT February 202686.4%
MathArena Apex10.9%
Kagi LLM Benchmark51.7%
ARC-AGI-197.5%44.7%
ARC-AGI-290.0%4.9%

Data as of 2026-09-10. Best configuration of each model; every score links to its source on the model pages.

Claude Fable 5.1 vs GLM-5: questions

Is Claude Fable 5.1 better than GLM-5?
Claude Fable 5.1 (max) leads on quality: 70.3 vs 59.1. The BenchLeader Index combines every independent quality benchmark; Claude Fable 5.1 (max) is ahead overall as of 2026-09-10, but check the category scores for your use.
Is Claude Fable 5.1 better than GLM-5 for coding?
Claude Fable 5.1 scores higher in coding (78 vs 53 on the category index, where 50 is average).
Is Claude Fable 5.1 better than GLM-5 for agentic tasks?
Claude Fable 5.1 scores higher in agentic tasks (67 vs 57 on the category index, where 50 is average).
Which is cheaper, Claude Fable 5.1 or GLM-5?
GLM-5 is cheaper: $1.55 against $20.00 per million tokens, blended at three input tokens per output token.
Which is faster, Claude Fable 5.1 or GLM-5?
GLM-5 streams faster: 66 against 40 output tokens per second.
Which has the larger context window?
Claude Fable 5.1 accepts more context: 1M against 205k tokens.