BenchLeader
Institute of Foundation ModelsOpen weightsReasoning modelNewFirst measured 18 Sept 2026

K2 Horizon 3.7B

K2 Horizon 3.7B is an Institute of Foundation Models open-weights reasoning model. It is not yet ranked: 3 independent results so far, and the index needs at least five across two categories.

Blended price
Output speed
First answer
Context
524k
Released
date not published by our sources
How it scores by categoryDashed line = average model (50). One step of 15 = one standard deviation.
  1. Knowledge51
  2. Long context57
  3. Composite49

Benchmark results

One column per reasoning effort. Rank is among every configuration of every model on that benchmark. Hover a score for the run it came from.

Reasoning

BenchmarkdefaultSource
GPQA Diamond (AA)not in index69.2%#302Artificial Analysis
Humanity's Last Exam (AA)not in index13.9%#213Artificial Analysis

Coding

BenchmarkdefaultSource
SciCode (AA)not in index22.0%#167Artificial Analysis

Knowledge

BenchmarkdefaultSource
AA-Omniscience-28.6#220Artificial Analysis

Long context

BenchmarkdefaultSource
AA-LCR62.3%#217Artificial Analysis

Composite

BenchmarkdefaultSource
AA Intelligence Index16.2#221Artificial Analysis

See also

Data as of 18 Sept 2026. Compare with another model.

Cite as: BenchLeader, “K2 Horizon 3.7B: benchmarks, pricing, speed and rank”, https://www.benchleader.com/models/k2-horizon-3-7b, data as of 18 Sept 2026.