BenchLeader

Gemini 3.6 Flash vs Muse Spark 1.3

Verdict
  • Muse Spark 1.3 (xhigh) leads on quality: 63.9 vs 60.8.
  • Gemini 3.6 Flash is stronger in maths, multimodal.
  • Muse Spark 1.3 (xhigh) is stronger in coding, composite, knowledge, long context, agents & tools.
  • Gemini 3.6 Flash is 1.3× cheaper ($1.50 vs $2.00 per 1M blended).
MetricGemini 3.6 FlashMuse Spark 1.3 (xhigh)
BenchLeader Index60.863.9
Coding score49.870.3
Composite score72.478.6
Knowledge score74.775.1
Long context score66.668.1
Maths score54.8
Multimodal score67.866.6
Agents & tools score60.2
Blended price $/M$1.50$2.00
Output speed194 tok/s190 tok/s
Time to first answer16.8 s35.7 s
Context window1.0M1M
FrontierCode34.4%
ProofBench36.0%
Epoch Capabilities Index154.3
LMArena WebDev1625
LiveBench81.6%
LiveBench Reasoning89.7%
LiveBench Coding81.1%
LiveBench Agentic Coding64.1%
LiveBench Mathematics96.0%
LiveBench Data Analysis79.6%
LiveBench Language82.8%
AA Intelligence Index34.345.2
AA-LCR80.0%83.0%
MMMU-Pro83.2%82.0%
AA-Omniscience22.123.1
GPQA Diamond (AA)92.8%94.1%
Humanity's Last Exam (AA)40.8%47.5%
SciCode (AA)53.4%59.7%
Terminal-Bench 2.1 (Vals)72.3%
Vals Index60.3
AIME 202696.7%
HMMT February 202689.4%
MathArena Apex26.0%

Data as of 2026-09-10. Best configuration of each model; every score links to its source on the model pages.

Gemini 3.6 Flash vs Muse Spark 1.3: questions

Is Gemini 3.6 Flash better than Muse Spark 1.3?
Muse Spark 1.3 (xhigh) leads on quality: 63.9 vs 60.8. The BenchLeader Index combines every independent quality benchmark; Muse Spark 1.3 (xhigh) is ahead overall as of 2026-09-10, but check the category scores for your use.
Is Gemini 3.6 Flash better than Muse Spark 1.3 for coding?
Muse Spark 1.3 scores higher in coding (70 vs 50 on the category index, where 50 is average).
Which is cheaper, Gemini 3.6 Flash or Muse Spark 1.3?
Gemini 3.6 Flash is cheaper: $1.50 against $2.00 per million tokens, blended at three input tokens per output token.
Which is faster, Gemini 3.6 Flash or Muse Spark 1.3?
Gemini 3.6 Flash streams faster: 194 against 190 output tokens per second.
Which has the larger context window?
Gemini 3.6 Flash accepts more context: 1.0M against 1M tokens.