BenchLeader
OpenAIAuto-detected

GPT 5.1 Instant

Not yet ranked: too few independent results so far.

Blended price
Output speed
First answer
Context
How it scores by categoryDashed line = average model (50). One step of 15 = one standard deviation.
  1. Knowledge35
  2. Instruction following43
  3. Multimodal39

Benchmark results

One column per reasoning effort. Rank is among every configuration of every model on that benchmark. Hover a score for the run it came from.

Knowledge

BenchmarkdefaultSource
MultiNRC18.3%#35Scale AI SEAL

Instruction following

BenchmarkdefaultSource
MultiChallenge51.2%#22Scale AI SEAL
TutorBench49.1%#20Scale AI SEAL

Multimodal

BenchmarkdefaultSource
VISTA34.9%#44Scale AI SEAL

See also

Data as of 9 Sept 2026. Compare with another model.