BenchLeader
AI21 LabsOpen weightsReasoning modelAuto-detected

Jamba Reasoning 3B

Best configuration ranks #463 of 610 on the BenchLeader Index at 42.3 ±7.3 (thinking reasoning effort).

Blended price
Output speed
First answer
Context
262k
How it scores by categoryDashed line = average model (50). One step of 15 = one standard deviation.
  1. Overall index42
  2. Agents & tools34
  3. Knowledge37
  4. Instruction following53
  5. Long context28
  6. Composite36

Benchmark results

One column per reasoning effort. Rank is among every configuration of every model on that benchmark. Hover a score for the run it came from.

Reasoning

BenchmarkthinkingSource
GPQA Diamond (AA)not in index33.3%#474Artificial Analysis
Humanity's Last Exam (AA)not in index3.8%#486Artificial Analysis

Agents & tools

Knowledge

BenchmarkthinkingSource
AA-Omniscience-57.7#372Artificial Analysis

Instruction following

BenchmarkthinkingSource
IFBench52.5%#158Artificial Analysis

Long context

BenchmarkthinkingSource
AA-LCR6.3%#409Artificial Analysis

Composite

BenchmarkthinkingSource
AA Intelligence Index5.7#478Artificial Analysis

See also

Data as of 9 Sept 2026. Compare with another model.