BenchLeader

DeepSeek V4 Pro vs Muse Spark 1.3

Verdict
  • Muse Spark 1.3 leads on quality: 68.3 vs 59.6.
  • DeepSeek V4 Pro (high) is stronger in agents & tools, instruction following.
  • Muse Spark 1.3 is stronger in composite, knowledge, long context.
  • DeepSeek V4 Pro (high) is 3.7× cheaper ($0.544 vs $2.00 per 1M blended).
  • Muse Spark 1.3 streams 3.2× faster (241 vs 75 tokens per second).
MetricDeepSeek V4 Pro (high)Muse Spark 1.3
BenchLeader Index59.668.3
Agents & tools score63.5
Composite score67.390.3
Instruction following score70.1
Knowledge score59.479.1
Long context score61.768.2
Blended price $/M$0.544$2.00
Output speed75 tok/s241 tok/s
Time to first answer28.2 s30.5 s
Context window1M1.0M
LMArena Agent4.1
AA Intelligence Index30.148.2
IFBench71.3%
AA-LCR70.3%83.0%
AA-Omniscience-10.625
Terminal-Bench Hard41.7%
GPQA Diamond (AA)90.5%93.5%
Humanity's Last Exam (AA)35.2%48.7%
SciCode (AA)58.8%
τ²-Bench Telecom (AA)94.2%
PRBench Finance59.5%
PRBench Legal61.6%

Data as of 2026-09-14. Best configuration of each model; every score links to its source on the model pages.

DeepSeek V4 Pro vs Muse Spark 1.3: questions

Is DeepSeek V4 Pro better than Muse Spark 1.3?
Muse Spark 1.3 leads on quality: 68.3 vs 59.6. The BenchLeader Index combines every independent quality benchmark; Muse Spark 1.3 is ahead overall as of 2026-09-14, but check the category scores for your use.
Which is cheaper, DeepSeek V4 Pro or Muse Spark 1.3?
DeepSeek V4 Pro is cheaper: $0.544 against $2.00 per million tokens, blended at three input tokens per output token.
Which is faster, DeepSeek V4 Pro or Muse Spark 1.3?
Muse Spark 1.3 streams faster: 241 against 75 output tokens per second.
Which has the larger context window?
Muse Spark 1.3 accepts more context: 1.0M against 1M tokens.