BenchLeader

Claude Fable 5.1 vs Qwen3.8 27B

Verdict
  • Claude Fable 5.1 (high) leads on quality: 72.0 vs 59.2.
  • Claude Fable 5.1 (high) is stronger in coding, composite, knowledge, long context, reasoning.
  • Qwen3.8 27B (xhigh) is stronger in agents & tools, multimodal.
  • Qwen3.8 27B (xhigh) is 21× cheaper ($0.938 vs $20.00 per 1M blended).
MetricClaude Fable 5.1 (high)Qwen3.8 27B (xhigh)
BenchLeader Index72.059.2
Agents & tools score66.667.3
Coding score74.153.2
Composite score94.072.0
Knowledge score83.756.4
Long context score68.467.5
Reasoning score81.052.5
Multimodal score60.7
Blended price $/M$20.00$0.938
Output speed55 tok/s42 tok/s
Time to first answer14.7 s51.0 s
Context window1M262k
Terminal-Bench54.5%
SciCode57.6%44.7%
WeirdML92.3%
APEX-Agents44.4%
AA Intelligence Index51.233.9
AA-LCR83.7%82.0%
MMMU-Pro76.3%
AA-Omniscience40.8-10.0
GPQA Diamond (AA)90.6%90.5%
Humanity's Last Exam (AA)55.9%33.9%
SciCode (AA)58.7%46.6%
LiveCodeBench84.0%
MMLU-Pro84.3%
IOI39.1%
LegalBench82.4%
TaxEval70.8%
Terminal-Bench 2.1 (Vals)58.4%
SWE-bench (Vals)86.0%
GPQA Diamond (Vals)88.9%
Vals Index48.5
ARC-AGI-196.0%
ARC-AGI-288.8%
CritPt30.3%5.4%
GDPval (AA)57.5%48.2%
τ²-Bench Banking (AA)43.1%48.0%
Code Migration14.2%
Excel Modeling Benchmark59.7%
Finance Agent v248.5%
Harvey's Legal Agent Benchmark11.3%
Legal Research Bench36.1%
MedCode28.7%
MedScribe83.8%
MMMU-Pro (Vals)83.9%
MortgageTax64.9%
ProgramBench0.0%
SAGE52.4%
SkillsBench38.1%
Terminal-Bench 4.0 (Vals)4.0%
Terminal-Bench Science1.4%
Vibe Code Bench v1.164.8%
MirrorCode73.3%
Surface Evolver Bench45.0%
CursorBench69.4%
ALE-Bench2143.2

Data as of 2026-09-19. Best configuration of each model; every score links to its source on the model pages.

Claude Fable 5.1 vs Qwen3.8 27B: questions

Is Claude Fable 5.1 better than Qwen3.8 27B?
Claude Fable 5.1 (high) leads on quality: 72.0 vs 59.2. The BenchLeader Index combines every independent quality benchmark; Claude Fable 5.1 (high) is ahead overall as of 2026-09-19, but check the category scores for your use.
Is Claude Fable 5.1 better than Qwen3.8 27B for coding?
Claude Fable 5.1 scores higher in coding (74 vs 53 on the category index, where 50 is average).
Is Claude Fable 5.1 better than Qwen3.8 27B for agentic tasks?
Qwen3.8 27B scores higher in agentic tasks (67 vs 67 on the category index, where 50 is average).
Which is cheaper, Claude Fable 5.1 or Qwen3.8 27B?
Qwen3.8 27B is cheaper: $0.938 against $20.00 per million tokens, blended at three input tokens per output token.
Which is faster, Claude Fable 5.1 or Qwen3.8 27B?
Claude Fable 5.1 streams faster: 55 against 42 output tokens per second.
Which has the larger context window?
Claude Fable 5.1 accepts more context: 1M against 262k tokens.