BenchLeader

LMArena Search

Pairwise votes on answers that used web search.

As of 19 Sept 2026, GPT-5.6 Sol leads LMArena Search on BenchLeader with 1256, ahead of Claude Fable 5 at 1229, across 33 model configurations with a published result.

Published by
LMArena
Category
Agents & tools
Index weight
Reference only
Models
33
Data as of
19 Sept 2026

CC BY 4.0 — lmarena-ai/leaderboard-dataset on Hugging Face.

What the test looks like

Search Arena compares models that answer with live web search, judged on the answer and its sourcing.

How it is scored

Bradley-Terry rating from pairwise votes, in Elo-like units, with style control.

What to keep in mind

Only search-enabled products are listed, and results depend on the day.

33 of 33
#
1GPT-5.6 SolxhighOpenAI125668.22026-08-24
2Claude Fable 5Anthropic122968.32026-08-24
3GPT 5.5 SearchOpenAI12242026-08-24
4Claude Opus 4.6 SearchAnthropic12232026-08-24
5Claude Opus 4.7Anthropic121264.52026-08-24
6Gemini 3.1 Pro GroundingGoogle12082026-08-24
7Grok 4.20 Multi-AgentxAI120559.22026-08-24
8Grok 4.5xAI120260.32026-08-24
9Gemini 3 Pro GroundingGoogle12012026-08-24
10Claude Sonnet 4.6 SearchAnthropic12012026-08-24
11Grok 4.20 Beta1xAI11972026-08-24
12Claude Sonnet 5 SearchAnthropic11962026-08-24
13Claude Opus 4.8Anthropic119561.92026-08-24
14GPT 5.4 SearchOpenAI11952026-08-24
15Grok 4.1 Fast SearchxAI11942026-08-24
16Ernie 5.1Baidu11932026-08-24
17Grok 4.3xAI119350.52026-08-24
18o3 SearchOpenAI11922026-08-24
19Gemini 3 Flash GroundingGoogle11912026-08-24
20Claude Opus 4.5 SearchAnthropic11872026-08-24
21GPT 5.1 SearchOpenAI11832026-08-24
22GPT 5 SearchOpenAI11832026-08-24
23Claude Sonnet 4.5 SearchAnthropic11752026-08-24
24Claude Opus 4.1 SearchAnthropic11672026-08-24
25GPT 5.2 SearchOpenAI11622026-08-24
26Grok 4 Fast SearchxAI11602026-08-24
27Claude Opus 4 SearchAnthropic11512026-08-24
28Ppl Sonar ProhighPerplexity11432026-08-24
29Gemini 2.5 Pro GroundingGoogle11422026-08-24
30GPT 5.2 Searchno reasoningOpenAI11372026-08-24
31Grok 4 SearchxAI11222026-08-24
32GPT 4o SearchOpenAI10812026-08-24
33Diffbot Small XlDiffbot10622026-08-24

Cite as: BenchLeader, “LMArena Search leaderboard”, https://www.benchleader.com/benchmarks/lmarena_search, data as of 19 Sept 2026.

LMArena Search: questions

What does LMArena Search measure?
Search Arena compares models that answer with live web search, judged on the answer and its sourcing. Scores are reported in Elo-style ratings from pairwise comparisons; higher is better.
Which AI model leads LMArena Search?
GPT-5.6 Sol leads LMArena Search with 1256 as of 19 Sept 2026, ahead of Claude Fable 5 at 1229.
How many models have LMArena Search results?
33 model configurations have a LMArena Search result on BenchLeader, all taken from LMArena.
Who runs LMArena Search and how often is it updated?
LMArena Search is published by LMArena. BenchLeader re-reads the published results every morning and records the date each result was published.
Does LMArena Search count toward the BenchLeader Index?
No. LMArena Search is shown for reference but left out of the composite index.