LMArena Search Arena
LMArena · search-2026-08-24-011508720696 · Blind human preferences for search-augmented responses in LMArena.
Model Ranking
Source | |||||||
|---|---|---|---|---|---|---|---|
| #1 | GPT-5.6 SolOpenAI | 1,257.26 | — | 29,663 | gpt-5.6-sol-xhigh; 95% CI [1249.79496224, 1264.72573796]; votes 29663; rank 1unknown | Original sourcewinner eligible | Third-party benchmark |
| #2 | Claude Fable 5Anthropic | 1,230.22 | — | 41,795 | claude-fable-5; 95% CI [1222.30895335, 1238.13609545]; votes 41795; rank 5unknown | Original sourcewinner eligible | Third-party benchmark |
| #3 | ERNIE 5.1Baidu | 1,226.96 | — | 3,788 | ernie-5.1; 95% CI [1217.24505738, 1236.67203321]; votes 3788; rank 6unknown | Original sourcewinner eligible | Third-party benchmark |
| #4 | Gemini 3.1 ProGoogle DeepMind | 1,210.46 | — | 113,282 | gemini-3.1-pro-grounding; 95% CI [1205.32063396, 1215.60669852]; votes 113282; rank 9unknown | Original sourcewinner eligible | Third-party benchmark |
| #5 | Claude Opus 4.8Anthropic | 1,204.30 | — | 70,998 | claude-opus-4-8; 95% CI [1197.89112684, 1210.71069451]; votes 70998; rank 12unknown | Original sourcewinner eligible | Third-party benchmark |
| #6 | Grok 4.20 Multi-AgentxAI | 1,204.11 | — | 109,553 | grok-4.20-multi-agent-beta-0309; 95% CI [1198.80874858, 1209.41102574]; votes 109553; rank 13unknown | Original sourcewinner eligible | Third-party benchmark |
| #7 | Gemini 3 FlashGoogle DeepMind | 1,198.09 | — | 149,334 | gemini-3-flash-grounding; 95% CI [1193.35426288, 1202.82591975]; votes 149334; rank 15unknown | Original sourcewinner eligible | Third-party benchmark |
| #8 | Gemini 2.5 ProGoogle DeepMind | 1,141.65 | — | 83,404 | gemini-2.5-pro-grounding; 95% CI [1137.05776179, 1146.24105969]; votes 83404; rank 27unknown | Original sourcewinner eligible | Third-party benchmark |
| #9 | Sonar Reasoning ProPerplexity | 1,138.60 | — | 29,055 | ppl-sonar-reasoning-pro-high; 95% CI [1132.86119454, 1144.34164089]; votes 29055; rank 29unknown | Original sourcewinner eligible | Third-party benchmark |
| #10 | Sonar ProPerplexitySelected model | 1,130.21 | — | 28,519 | ppl-sonar-pro-high; 95% CI [1124.48384666, 1135.92956299]; votes 28519; rank 31unknown | Original sourcewinner eligible | Third-party benchmark |
10 models ranked by the best winner-eligible current-version score when available, otherwise the best publisher-artifact score; highest first. Observed Sep 2, 2026. The model you came from is highlighted.
Methodology and Coverage
Primary Evidence
Publisher Artifacts
Questions
LMArena Search Arena FAQs
What does LMArena Search Arena measure?+
Blind human preferences for search-augmented responses in LMArena. Model Markets classifies it as a search benchmark and preserves the publisher's search-2026-08-24-011508720696 release as a distinct comparison cohort.
How are models ranked on LMArena Search Arena?+
Models are ordered by arena rating in rating, with higher scores ranked first. Each model appears once; a verified winner-eligible current-version result takes precedence over a publisher-artifact-only result, and missing scores are not estimated.
Which model currently leads LMArena Search Arena?+
GPT-5.6 Sol leads the current verified table with 1,257.26 rating on search-2026-08-24-011508720696. This is a benchmark-specific result, not a universal model-quality claim.
How many models have a published LMArena Search Arena score?+
10 catalog models are published from 34 source rows. 24 source identities remain quarantined rather than guessed.
Can LMArena Search Arena scores be compared with other benchmarks?+
Raw scores should be compared only within the same benchmark version, metric, and protocol. Model Markets normalizes eligible scores only for aggregate rankings, and groups models by an identical benchmark set before ranking them. View aggregate rankings →
Why might a model be missing from LMArena Search Arena?+
A model remains absent when the publisher has no current result, the source model identity is unresolved, the evaluation protocol is incompatible, or the evidence cannot be verified. Model Markets does not substitute a provider claim or infer a score from a related model.
How current is the LMArena Search Arena leaderboard?+
The current Model Markets snapshot was observed Sep 2, 2026 from LMArena artifacts. The exact publisher source and each retained result artifact are linked on this page.