BenchLeader

GLM 5 Turbo vs Muse Spark 1.1

Verdict
  • Muse Spark 1.1 leads on quality: 65.0 vs 58.4.
  • GLM 5 Turbo is stronger in agents & tools.
  • Muse Spark 1.1 is stronger in composite, instruction following, knowledge, long context, coding, human preference, maths, multimodal, reasoning.
  • They cost about the same ($1.90 per 1M blended).
  • Muse Spark 1.1 streams 9.9× faster (179 vs 18 tokens per second).
MetricGLM 5 TurboMuse Spark 1.1
BenchLeader Index58.465.0
Agents & tools score63.162.3
Composite score62.972.6
Instruction following score71.877.8
Knowledge score56.673.3
Long context score62.465.5
Coding score68.0
Human preference score66.4
Maths score52.2
Multimodal score65.0
Reasoning score69.3
Blended price $/M$1.90$2.00
Output speed18 tok/s179 tok/s
Time to first answer7.9 s2.3 s
Context window200k1.0M
SimpleQA Verified57.8%
SciCode58.2%
APEX-Agents41.9%
ProofBench39.0%
Epoch Capabilities Index154.6
LMArena Text1494
LMArena Hard Prompts1513
LMArena Coding1533
LMArena WebDev1542
LMArena Vision1293
LMArena Agent-3
AA Intelligence Index26.634.3
IFBench73.2%
AA-LCR71.7%77.7%
AA-Omniscience-16.428.1
Terminal-Bench Hard33.3%
GPQA Diamond (AA)84.8%89.8%
Humanity's Last Exam (AA)27.8%46.2%
SciCode (AA)58.8%
τ²-Bench Telecom (AA)98.5%
SWE-Bench Pro61.5%
MCP Atlas88.1%
MultiChallenge75.3%
PRBench Finance55.0%
PRBench Legal57.0%
MultiNRC65.6%
EQ-Bench 41260

Data as of 2026-09-13. Best configuration of each model; every score links to its source on the model pages.

GLM 5 Turbo vs Muse Spark 1.1: questions

Is GLM 5 Turbo better than Muse Spark 1.1?
Muse Spark 1.1 leads on quality: 65.0 vs 58.4. The BenchLeader Index combines every independent quality benchmark; Muse Spark 1.1 is ahead overall as of 2026-09-13, but check the category scores for your use.
Is GLM 5 Turbo better than Muse Spark 1.1 for agentic tasks?
GLM 5 Turbo scores higher in agentic tasks (63 vs 62 on the category index, where 50 is average).
Which is cheaper, GLM 5 Turbo or Muse Spark 1.1?
GLM 5 Turbo is cheaper: $1.90 against $2.00 per million tokens, blended at three input tokens per output token.
Which is faster, GLM 5 Turbo or Muse Spark 1.1?
Muse Spark 1.1 streams faster: 179 against 18 output tokens per second.
Which has the larger context window?
Muse Spark 1.1 accepts more context: 1.0M against 200k tokens.