BenchLeader

GPT-5 Pro vs Grok 4.3

Verdict
  • GPT-5 Pro and Grok 4.3 (medium) are level on quality (60.8 vs 60.7).
  • GPT-5 Pro is stronger in multimodal, reasoning.
  • Grok 4.3 (medium) is stronger in knowledge, agents & tools, composite, instruction following, long context.
  • Grok 4.3 (medium) is 26× cheaper ($1.56 vs $41.25 per 1M blended).
MetricGPT-5 ProGrok 4.3 (medium)
BenchLeader Index60.860.7
Knowledge score68.072.1
Multimodal score67.460.2
Reasoning score57.6
Agents & tools score60.1
Composite score60.3
Instruction following score80.0
Long context score64.0
Blended price $/M$41.25$1.56
Output speed112 tok/s
Time to first answer12.1 s
Context window400k1M
Humanity's Last Exam31.6%
SimpleBench61.6%
Epoch Capabilities Index150.3
AA Intelligence Index24.8
IFBench83.3%
AA-LCR75.0%
MMMU-Pro75.8%
AA-Omniscience16.7
Terminal-Bench Hard30.3%
GPQA Diamond (AA)89.0%
Humanity's Last Exam (AA)30.0%
τ²-Bench Telecom (AA)91.2%
PRBench Finance51.1%
PRBench Legal49.9%
VISTA52.4%
MultiNRC65.2%
Kagi LLM Benchmark76.8%
ARC-AGI-170.2%
ARC-AGI-218.3%

Data as of 2026-09-10. Best configuration of each model; every score links to its source on the model pages.

GPT-5 Pro vs Grok 4.3: questions

Is GPT-5 Pro better than Grok 4.3?
GPT-5 Pro and Grok 4.3 (medium) are level on quality (60.8 vs 60.7). The BenchLeader Index combines every independent quality benchmark; GPT-5 Pro is ahead overall as of 2026-09-10, but check the category scores for your use.
Which is cheaper, GPT-5 Pro or Grok 4.3?
Grok 4.3 is cheaper: $1.56 against $41.25 per million tokens, blended at three input tokens per output token.
Which has the larger context window?
Grok 4.3 accepts more context: 1M against 400k tokens.