LMArena · Image understanding (vision)
We regularly fetch and display public benchmark data: LMArena user-voted rankings and Epoch AI test scores. Each leaderboard shows its publication and retrieval dates. These scores are not produced by this site.
How to read this leaderboard
What it measures: An image understanding ranking based on votes comparing two models' answers to questions that include images.
How to read the score: A relative score reflecting which answers users preferred. It is not an objective accuracy rate for interpreting photos, charts or screenshots.
Caveats: Check accuracy separately for the conditions you need, such as reading small text in documents or recognizing Korean (Hangul) characters.
Insights from this ranking
- The 95% confidence intervals for 1st-place claude-fable-5-high and 2nd-place claude-opus-5-high overlap. This aggregation alone does not clearly establish their order.
- 11 other models have confidence intervals that overlap with the 1st-place model's. Interpret small ranking differences cautiously alongside vote counts.
- Anthropic has the most models among the top 10, with 7.
- The 1st-place model received 13,709 votes, and this leaderboard includes 50 models.
Image understanding (vision) LMArena
Published 2026-10-02 · Retrieved 2026-10-04
Dots show scores; horizontal lines show 95% confidence intervals. Overlapping intervals suggest similar performance.
View table (top 50)
| Rank | Model | Developer | Score | 95% confidence interval | Votes |
|---|---|---|---|---|---|
| 1 | claude-fable-5-high | Anthropic | 1325 | 1318–1332 | 13,709 |
| 2 | claude-opus-5-high | Anthropic | 1321 | 1314–1328 | 15,900 |
| 3 | claude-fable-5.1-max | Anthropic | 1320 | 1310–1330 | 4,312 |
| 4 | claude-opus-4-7 | Anthropic | 1316 | 1310–1323 | 23,169 |
| 5 | gemini-3.8-flash-high | 1316 | 1304–1327 | 3,086 | |
| 6 | claude-opus-4-6-high | Anthropic | 1316 | 1309–1322 | 22,328 |
| 7 | gemini-3.7-flash-high | 1315 | 1305–1325 | 3,814 | |
| 8 | qwen3.8-max | Alibaba | 1314 | 1307–1322 | 10,040 |
| 9 | claude-opus-4-6 | Anthropic | 1313 | 1307–1319 | 26,869 |
| 10 | claude-opus-4-7-high | Anthropic | 1313 | 1306–1319 | 22,755 |
| 11 | muse-spark-1.3-max | Meta | 1310 | 1300–1321 | 3,843 |
| 12 | gemini-3.5-flash-high | 1310 | 1303–1317 | 12,336 | |
| 13 | gemini-3.5-flash-medium | 1308 | 1301–1315 | 12,467 | |
| 14 | muse-spark | Meta | 1305 | 1296–1314 | 5,868 |
| 15 | gemini-3-pro | 1305 | 1298–1312 | 13,620 | |
| 16 | muse-spark-1.2 (xHigh) | Meta | 1305 | 1290–1319 | 1,951 |
| 17 | Ogpt-5.4-high | OpenAI | 1302 | 1296–1308 | 29,177 |
| 18 | Zglm-5.3-flash | Z.ai | 1301 | 1292–1310 | 5,479 |
| 19 | Ogpt-5.5 | OpenAI | 1297 | 1291–1303 | 27,223 |
| 20 | gemini-3.6-flash-high | 1297 | 1288–1305 | 7,392 | |
| 21 | gemini-3.1-pro-preview | 1296 | 1291–1301 | 45,558 | |
| 22 | claude-opus-4-8-high | Anthropic | 1295 | 1288–1301 | 18,350 |
| 23 | muse-spark-1.1 | Meta | 1294 | 1286–1301 | 11,255 |
| 24 | Ogpt-5.5-high | OpenAI | 1292 | 1286–1298 | 25,488 |
| 25 | Ogpt-5.4 | OpenAI | 1292 | 1286–1299 | 22,828 |
| 26 | claude-sonnet-5.5-xhigh | Anthropic | 1290 | 1274–1305 | 1,608 |
| 27 | claude-opus-4-8 | Anthropic | 1289 | 1282–1296 | 18,857 |
| 28 | xgrok-4.5 | xAI | 1288 | 1280–1295 | 11,039 |
| 29 | gemini-3-flash | 1285 | 1280–1291 | 39,142 | |
| 30 | Ogpt-6.1-sol-max | OpenAI | 1285 | 1269–1301 | 1,408 |
| 31 | claude-sonnet-4-6 | Anthropic | 1283 | 1276–1289 | 27,429 |
| 32 | kimi-k2.6 | Moonshot AI | 1282 | 1276–1289 | 16,387 |
| 33 | qwen3.7-plus | Alibaba | 1280 | 1273–1287 | 13,065 |
| 34 | Ogpt-6-astra-max | OpenAI | 1280 | 1268–1291 | 3,175 |
| 35 | Ogpt-5.6-sol-xhigh | OpenAI | 1279 | 1272–1287 | 10,421 |
| 36 | SStep 5 Preview | Stepfun | 1278 | 1254–1302 | 604 |
| 37 | gemma-4-31b | 1277 | 1271–1283 | 39,111 | |
| 38 | claude-sonnet-5-high | Anthropic | 1275 | 1268–1282 | 13,029 |
| 39 | dola-seed-2.0-pro | ByteDance | 1275 | 1267–1282 | 10,981 |
| 40 | Ogpt-5.6-terra-xhigh | OpenAI | 1271 | 1264–1279 | 10,292 |
| 41 | qwen3.8-27b | Alibaba | 1270 | 1261–1279 | 6,422 |
| 42 | kimi-k2.5-thinking | Moonshot AI | 1269 | 1263–1275 | 28,326 |
| 43 | Ogpt-5.2-chat-latest-20260210 | OpenAI | 1268 | 1261–1274 | 16,480 |
| 44 | gemini-3.5-flash-lite | 1266 | 1258–1275 | 7,579 | |
| 45 | gemini-3-flash (thinking-minimal) | 1266 | 1260–1271 | 37,184 | |
| 46 | xgrok-4.6-high | xAI | 1265 | 1256–1274 | 5,736 |
| 47 | Zglm-5v-turbo | Z.ai | 1264 | 1259–1269 | 41,533 |
| 48 | qwen3.5-397b-a17b | Alibaba | 1264 | 1258–1269 | 31,441 |
| 49 | mimo-v2.6-pro | Xiaomi | 1264 | 1249–1278 | 1,787 |
| 50 | gemini-2.5-pro | 1263 | 1259–1268 | 88,813 |
Data: LMArena leaderboard dataset (CC BY 4.0). Changes: Top 50 entries per leaderboard; scores are rounded. https://huggingface.co/datasets/lmarena-ai/leaderboard-datasetCompany logos are trademarks of their respective owners and are used only for identification (icons: Simple Icons).Data: Epoch AI — Capabilities & benchmarking (CC BY 4.0). https://epoch.ai/benchmarks