BenchLeader
PoolsideOpen weightsReasoning modelAuto-detected

Laguna M.1

Best configuration ranks #560 of 610 on the BenchLeader Index at 37.5 ±9.7. Last measured 8 Sept 2026. Released 28 Apr 2026.

Blended price
Output speed
First answer
Context
262k
How it scores by categoryDashed line = average model (50). One step of 15 = one standard deviation.
  1. Overall index38
  2. Coding44
  3. Agents & tools24
  4. Maths31
  5. Knowledge32

Benchmark results

One column per reasoning effort. Rank is among every configuration of every model on that benchmark. Hover a score for the run it came from.

Reasoning

BenchmarkdefaultSourceTrend
GPQA Diamond (Vals)not in index27.0%#129Vals AI

Coding

BenchmarkdefaultSourceTrend
LMArena WebDev1347#91LMArena
LiveCodeBench68.1%#90Vals AI
SWE-bench (Vals)not in index57.6%#76Vals AI

Agents & tools

BenchmarkdefaultSourceTrend
Terminal-Bench 2.1 (Vals)34.1%#58Vals AI

Maths

BenchmarkdefaultSourceTrend
ProofBench0.0%#60Vals AI

Knowledge

BenchmarkdefaultSourceTrend
MMLU-Pro68.8%#119Vals AI
LegalBench75.1%#106Vals AI
CorpFin58.2%#83Vals AI
TaxEval1.6%#137Vals AI

See also

Data as of 9 Sept 2026. Compare with another model.