BenchLeader
Trillion LabsOpen weightsReasoning modelAuto-detected

Tri-21B-Think

Best configuration ranks #427 of 610 on the BenchLeader Index at 44.0 ±6.7 (thinking reasoning effort).

Blended price
Output speed
First answer
Context
32k
How it scores by categoryDashed line = average model (50). One step of 15 = one standard deviation.
  1. Overall index44
  2. Agents & tools36
  3. Knowledge35
  4. Instruction following55
  5. Long context36
  6. Composite41

Benchmark results

One column per reasoning effort. Rank is among every configuration of every model on that benchmark. Hover a score for the run it came from.

Reasoning

BenchmarkthinkingSource
GPQA Diamond (AA)not in index60.1%#357Artificial Analysis
Humanity's Last Exam (AA)not in index5.9%#336Artificial Analysis

Agents & tools

Knowledge

BenchmarkthinkingSource
AA-Omniscience-62.3#402Artificial Analysis

Instruction following

BenchmarkthinkingSource
IFBench54.6%#148Artificial Analysis

Long context

BenchmarkthinkingSource
AA-LCR21.7%#353Artificial Analysis

Composite

BenchmarkthinkingSource
AA Intelligence Index9.6#323Artificial Analysis

See also

Data as of 9 Sept 2026. Compare with another model.