BenchLeader
xAIAuto-detected

Grok 2

Best configuration ranks #480 of 610 on the BenchLeader Index at 41.6 ±4.9. Last measured 2 Sept 2026. Released 12 Dec 2024.

Blended price
Output speed
First answer
Context
131k
How it scores by categoryDashed line = average model (50). One step of 15 = one standard deviation.
  1. Overall index42
  2. Reasoning37
  3. Coding32
  4. Maths38
  5. Knowledge45
  6. Human preference51
  7. Composite38

Benchmark results

One column per reasoning effort. Rank is among every configuration of every model on that benchmark. Hover a score for the run it came from.

Coding

BenchmarkdefaultSource
WeirdML22.2%#133WeirdML
LMArena Coding1358#211LMArena
LiveCodeBench38.7%#125Vals AI

Knowledge

BenchmarkdefaultSource
MMLU-Pro75.5%#104Vals AI
CorpFin51.1%#99Vals AI
TaxEval67.0%#106Vals AI
MedQA92.3%#32Vals AI

Human preference

BenchmarkdefaultSource
LMArena Text1335#197LMArena

See also

Data as of 9 Sept 2026. Compare with another model.