LMArena Vision Arena
LMArena · vision-2026-08-27-011508720696 · Blind human preferences for multimodal responses in LMArena's vision arena.
Model Ranking
Source | |||||||
|---|---|---|---|---|---|---|---|
| #1 | Claude Fable 5Anthropic | 1,329.46 | — | 10,002 | claude-fable-5; 95% CI [1321.03632319, 1337.88661881]; votes 10002; rank 1unknown | Original sourcewinner eligible | Third-party benchmark |
| #2 | Qwen3.8-MaxQwen | 1,312.86 | — | 7,244 | qwen3.8-max; 95% CI [1304.57707433, 1321.13715544]; votes 7244; rank 6unknown | Original sourcewinner eligible | Third-party benchmark |
| #3 | Gemini 3.5 FlashGoogle DeepMind | 1,310.66 | — | 8,259 | gemini-3.5-flash-high; 95% CI [1302.35646187, 1318.96740220]; votes 8259; rank 8unknown | Original sourcewinner eligible | Third-party benchmark |
| #4 | Gemini 3.6 FlashGoogle DeepMind | 1,301.97 | — | 3,816 | gemini-3.6-flash-high; 95% CI [1291.15288726, 1312.79182351]; votes 3816; rank 13unknown | Original sourcewinner eligible | Third-party benchmark |
| #5 | GLM-5.3-FlashZ.ai | 1,296.37 | — | 1,389 | glm-5.3-flash; 95% CI [1279.58251138, 1313.15015420]; votes 1389; rank 15unknown | Original sourcewinner eligible | Third-party benchmark |
| #6 | Gemini 3.1 ProGoogle DeepMind | 1,295.39 | — | 39,412 | gemini-3.1-pro-preview; 95% CI [1289.75236576, 1301.03738662]; votes 39412; rank 17unknown | Original sourcewinner eligible | Third-party benchmark |
| #7 | Claude Opus 4.8Anthropic | 1,293.93 | — | 13,521 | claude-opus-4-8-high; 95% CI [1286.47351506, 1301.38306277]; votes 13521; rank 19unknown | Original sourcewinner eligible | Third-party benchmark |
| #8 | Gemini 3 FlashGoogle DeepMind | 1,284.99 | — | 36,747 | gemini-3-flash; 95% CI [1279.54770336, 1290.43069010]; votes 36747; rank 24unknown | Original sourcewinner eligible | Third-party benchmark |
| #9 | Qwen3.7-PlusQwenSelected model | 1,281.06 | — | 9,132 | qwen3.7-plus; 95% CI [1273.03498970, 1289.09217677]; votes 9132; rank 26unknown | Original sourcewinner eligible | Third-party benchmark |
| #10 | Kimi-K2.6Moonshot AI | 1,280.71 | — | 15,378 | kimi-k2.6; 95% CI [1273.64165646, 1287.78511818]; votes 15378; rank 27unknown | Original sourcewinner eligible | Third-party benchmark |
| #11 | Qwen3.8-27BQwen | 1,279.32 | — | 2,954 | qwen3.8-27b; 95% CI [1267.46655532, 1291.16933290]; votes 2954; rank 28unknown | Original sourcewinner eligible | Third-party benchmark |
| #12 | Claude Sonnet 5Anthropic | 1,277.94 | — | 8,522 | claude-sonnet-5-high; 95% CI [1269.73992861, 1286.14906472]; votes 8522; rank 29unknown | Original sourcewinner eligible | Third-party benchmark |
| #13 | GPT-5.6 SolOpenAI | 1,276.90 | — | 5,837 | gpt-5.6-sol-xhigh; 95% CI [1267.71657009, 1286.08047215]; votes 5837; rank 30unknown | Original sourcewinner eligible | Third-party benchmark |
| #14 | GPT-5.6 TerraOpenAI | 1,268.99 | — | 5,774 | gpt-5.6-terra-xhigh; 95% CI [1259.82673197, 1278.15538857]; votes 5774; rank 33unknown | Original sourcewinner eligible | Third-party benchmark |
| #15 | Gemini 3.5 Flash-LiteGoogle DeepMind | 1,268.89 | — | 3,886 | gemini-3.5-flash-lite; 95% CI [1258.10808094, 1279.66830878]; votes 3886; rank 34unknown | Original sourcewinner eligible | Third-party benchmark |
| #16 | Kimi-K2.5Moonshot AI | 1,266.74 | — | 26,728 | kimi-k2.5-thinking; 95% CI [1260.83834046, 1272.64651923]; votes 26728; rank 36unknown | Original sourcewinner eligible | Third-party benchmark |
| #17 | Qwen3.5-397B-A17BQwen | 1,264.80 | — | 25,997 | qwen3.5-397b-a17b; 95% CI [1258.85111028, 1270.75814841]; votes 25997; rank 38unknown | Original sourcewinner eligible | Third-party benchmark |
| #18 | Gemini 2.5 ProGoogle DeepMind | 1,262.42 | — | 87,719 | gemini-2.5-pro; 95% CI [1257.72056767, 1267.11330698]; votes 87719; rank 41unknown | Original sourcewinner eligible | Third-party benchmark |
| #19 | Grok 4.20 Multi-AgentxAI | 1,259.97 | — | 22,988 | grok-4.20-multi-agent-beta-0309; 95% CI [1253.56879963, 1266.36768970]; votes 22988; rank 42unknown | Original sourcewinner eligible | Third-party benchmark |
| #20 | MiniMax-M3MiniMax | 1,255.42 | — | 13,390 | minimax-m3; 95% CI [1247.93588333, 1262.90862002]; votes 13390; rank 44unknown | Original sourcewinner eligible | Third-party benchmark |
| #21 | GPT-5.6 LunaOpenAI | 1,254.30 | — | 5,857 | gpt-5.6-luna-xhigh; 95% CI [1245.05722338, 1263.53552757]; votes 5857; rank 45unknown | Original sourcewinner eligible | Third-party benchmark |
| #22 | Gemini 2.5 FlashGoogle DeepMind | 1,235.76 | — | 58,968 | gemini-2.5-flash; 95% CI [1231.05290901, 1240.46508391]; votes 58968; rank 59unknown | Original sourcewinner eligible | Third-party benchmark |
| #23 | Mistral Large 3Mistral AI | 1,226.84 | — | 4,751 | mistral-large-3; 95% CI [1216.90651393, 1236.78308197]; votes 4751; rank 66unknown | Original sourcewinner eligible | Third-party benchmark |
| #24 | Mistral-Medium-3.5-128BMistral AI | 1,222.18 | — | 5,091 | mistral-medium-3.5; 95% CI [1212.39755732, 1231.95907443]; votes 5091; rank 67unknown | Original sourcewinner eligible | Third-party benchmark |
| #25 | GPT-4.1OpenAI | 1,210.32 | — | 41,034 | gpt-4.1-2025-04-14; 95% CI [1203.63813827, 1217.00696914]; votes 41034; rank 70unknown | Original sourcewinner eligible | Third-party benchmark |
| #26 | GPT-4.1 MiniOpenAI | 1,181.17 | — | 40,292 | gpt-4.1-mini-2025-04-14; 95% CI [1173.69296987, 1188.64000739]; votes 40292; rank 84unknown | Original sourcewinner eligible | Third-party benchmark |
| #27 | Llama-4-Maverick-17B-128E-InstructMeta | 1,141.47 | — | 6,936 | llama-4-maverick-17b-128e-instruct; 95% CI [1132.42029588, 1150.51619722]; votes 6936; rank 102unknown | Original sourcewinner eligible | Third-party benchmark |
| #28 | Llama-4-Scout-17B-16E-InstructMeta | 1,118.08 | — | 6,467 | llama-4-scout-17b-16e-instruct; 95% CI [1108.62969338, 1127.52426769]; votes 6467; rank 110unknown | Original sourcewinner eligible | Third-party benchmark |
| #29 | GPT-4o MiniOpenAI | 1,065.63 | — | 17,341 | gpt-4o-mini-2024-07-18; 95% CI [1057.40637002, 1073.85507682]; votes 17341; rank 118unknown | Original sourcewinner eligible | Third-party benchmark |
| #30 | GPT-4oOpenAI | 1,064.54 | — | 3,376 | gpt-4o-2024-08-06; 95% CI [1052.10281463, 1076.97369812]; votes 3376; rank 119unknown | Original sourcewinner eligible | Third-party benchmark |
| #31 | GPT-4.1 NanoOpenAI | 1,063.23 | — | 1,211 | gpt-4.1-nano-2025-04-14; 95% CI [1044.95116252, 1081.50991534]; votes 1211; rank 120unknown | Original sourcewinner eligible | Third-party benchmark |
31 models ranked by the best winner-eligible current-version score when available, otherwise the best publisher-artifact score; highest first. Observed Sep 2, 2026. The model you came from is highlighted.
Methodology and Coverage
Primary Evidence
Publisher Artifacts
Questions
LMArena Vision Arena FAQs
What does LMArena Vision Arena measure?+
Blind human preferences for multimodal responses in LMArena's vision arena. Model Markets classifies it as a multimodal benchmark and preserves the publisher's vision-2026-08-27-011508720696 release as a distinct comparison cohort.
How are models ranked on LMArena Vision Arena?+
Models are ordered by arena rating in rating, with higher scores ranked first. Each model appears once; a verified winner-eligible current-version result takes precedence over a publisher-artifact-only result, and missing scores are not estimated.
Which model currently leads LMArena Vision Arena?+
Claude Fable 5 leads the current verified table with 1,329.46 rating on vision-2026-08-27-011508720696. This is a benchmark-specific result, not a universal model-quality claim.
How many models have a published LMArena Vision Arena score?+
31 catalog models are published from 148 source rows. 114 source identities remain quarantined rather than guessed.
Can LMArena Vision Arena scores be compared with other benchmarks?+
Raw scores should be compared only within the same benchmark version, metric, and protocol. Model Markets normalizes eligible scores only for aggregate rankings, and groups models by an identical benchmark set before ranking them. View aggregate rankings →
Why might a model be missing from LMArena Vision Arena?+
A model remains absent when the publisher has no current result, the source model identity is unresolved, the evaluation protocol is incompatible, or the evidence cannot be verified. Model Markets does not substitute a provider claim or infer a score from a related model.
How current is the LMArena Vision Arena leaderboard?+
The current Model Markets snapshot was observed Sep 2, 2026 from LMArena artifacts. The exact publisher source and each retained result artifact are linked on this page.