LMArena · Japanese
We regularly fetch and display public benchmark data: LMArena user voting rankings, Epoch AI test scores, and OpenRouter usage rankings. Each leaderboard shows both its publication date and the date we retrieved it. These scores are not assigned by this site.
How to read this leaderboard
What it measures: A ranking calculated from votes on questions classified as Japanese in LMArena text conversations.
How to read the score: Shows which models voters preferred for questions in Japanese. The question's language does not indicate the voter's native language or identity.
Caveats: Vote counts and confidence interval widths vary by language and model. For new models, check vote counts, update dates and aggregation conditions together.
Insights from this ranking
- The 95% confidence intervals for 1st-place claude-fable-5.1-max and 2nd-place claude-fable-5-high overlap. This aggregation alone does not clearly establish their order.
- 25 other models have confidence intervals that overlap with the 1st-place model's. Interpret small ranking differences cautiously alongside vote counts.
- Google has the most models among the top 10, with 4.
- The 1st-place model received 200 votes, and this leaderboard includes 274 models.
Japanese LMArena
Published 2026-10-02 · Retrieved 2026-10-04
Dots show scores; horizontal lines show 95% confidence intervals. Overlapping intervals suggest similar performance. The chart shows only the highest-scoring reasoning setting, such as high or max, for each model. The table shows rankings for all settings. Chat and vision rankings use 'style and length adjustment (style control)', the same default as the LMArena website.
Top ranking by company
- Anthropicclaude-fable-5.1-max#1
- Googlegemini-3-pro#3
- OOpenAIgpt-5.5-high#6
- Moonshot AIkimi-k3-max#8
- Metamuse-spark-1.3-max#10
- Alibabaqwen3.5-max-preview#15
- xxAIgrok-4.20-beta1#20
- DeepSeekdeepseek-v4-pro#31
- ZZ.aiglm-5.2-max#33
View table (top 50)
| Rank | Model | Developer | Score | 95% confidence interval | Votes |
|---|---|---|---|---|---|
| 1 | #1claude-fable-5.1-max | Anthropic | 1542 | 1496–1588 | 200 |
| 2 | #2claude-fable-5-high | Anthropic | 1524 | 1496–1551 | 538 |
| 3 | #3gemini-3-pro | 1508 | 1478–1538 | 471 | |
| 4 | #4claude-opus-5-high | Anthropic | 1506 | 1483–1528 | 958 |
| 5 | #5gemini-3.7-flash-high | 1504 | 1471–1538 | 340 | |
| 6 | #6Ogpt-5.5-high | OpenAI | 1503 | 1481–1525 | 895 |
| 7 | #7gemini-3.8-flash-high | 1503 | 1473–1533 | 462 | |
| 8 | #8kimi-k3-max | Moonshot AI | 1502 | 1474–1530 | 508 |
| 9 | #9gemini-3.1-pro-preview | 1499 | 1481–1517 | 1,409 | |
| 10 | #10muse-spark-1.3-max | Meta | 1498 | 1452–1544 | 188 |
| 11 | #11claude-opus-5-max | Anthropic | 1497 | 1467–1526 | 482 |
| 12 | #12claude-opus-4-6-high | Anthropic | 1495 | 1473–1518 | 845 |
| 13 | #13Ogpt-5.6-sol-xhigh | OpenAI | 1495 | 1468–1523 | 549 |
| 14 | #14gemini-3-flash | 1487 | 1453–1521 | 337 | |
| 15 | #15qwen3.5-max-preview | Alibaba | 1486 | 1442–1531 | 207 |
| 16 | #16claude-opus-4-7-high | Anthropic | 1485 | 1461–1510 | 705 |
| 17 | #17claude-opus-4-6 | Anthropic | 1480 | 1459–1502 | 875 |
| 18 | #18gemini-3.5-flash-high | 1479 | 1454–1504 | 672 | |
| 19 | #19Ogpt-5.4-high | OpenAI | 1479 | 1454–1504 | 671 |
| 20 | #20xgrok-4.20-beta1 | xAI | 1478 | 1437–1518 | 225 |
| 21 | #21gemini-3.6-flash-high | 1478 | 1452–1504 | 598 | |
| 22 | #22Ogpt-5.5 | OpenAI | 1477 | 1454–1501 | 800 |
| 23 | #23qwen3.8-max | Alibaba | 1473 | 1442–1505 | 402 |
| 24 | #24claude-opus-4-7 | Anthropic | 1473 | 1450–1496 | 797 |
| 25 | #25muse-spark-1.1 | Meta | 1471 | 1445–1496 | 633 |
| 26 | #26Ogpt-5.6-terra-xhigh | OpenAI | 1466 | 1439–1494 | 559 |
| 27 | #27gemini-3.5-flash-medium | 1461 | 1436–1486 | 668 | |
| 28 | #28Ogpt-5.1-high | OpenAI | 1460 | 1432–1488 | 458 |
| 29 | #29Ogpt-5.5-instant | OpenAI | 1460 | 1422–1498 | 258 |
| 30 | #30claude-opus-4-8 | Anthropic | 1458 | 1436–1480 | 874 |
| 31 | #31deepseek-v4-pro | DeepSeek | 1457 | 1433–1481 | 715 |
| 32 | #32claude-opus-4-8-high | Anthropic | 1456 | 1434–1478 | 881 |
| 33 | #33Zglm-5.2-max | Z.ai | 1455 | 1431–1480 | 702 |
| 34 | #34Zglm-5.3-max | Z.ai | 1454 | 1420–1488 | 327 |
| 35 | #35Ogpt-5.2-chat-latest-20260210 | OpenAI | 1449 | 1412–1485 | 266 |
| 36 | #36gemini-3.5-flash-lite | 1446 | 1419–1473 | 565 | |
| 37 | #37gemini-2.5-pro | 1445 | 1430–1460 | 2,037 | |
| 38 | #38Ogpt-5.4 | OpenAI | 1445 | 1421–1469 | 680 |
| 39 | #39claude-opus-4-5-20251101-high-32k | Anthropic | 1444 | 1413–1476 | 376 |
| 40 | #40claude-opus-4-5-20251101 | Anthropic | 1441 | 1417–1466 | 633 |
| 41 | #41kimi-k2.6 | Moonshot AI | 1441 | 1413–1469 | 497 |
| 42 | #42claude-sonnet-5-high | Anthropic | 1441 | 1416–1466 | 662 |
| 43 | #43xgrok-4.5 | xAI | 1440 | 1412–1468 | 522 |
| 44 | #44Zglm-4.7 | Z.ai | 1440 | 1385–1494 | 123 |
| 45 | #45claude-sonnet-4-6 | Anthropic | 1440 | 1417–1462 | 792 |
| 46 | #46Ogpt-4.5-preview-2025-02-27 | OpenAI | 1439 | 1401–1477 | 271 |
| 47 | #47Ogpt-5.2-high | OpenAI | 1437 | 1407–1467 | 443 |
| 48 | #48Zglm-5.1 | Z.ai | 1436 | 1412–1459 | 754 |
| 49 | #49gemini-3-flash (thinking-minimal) | 1434 | 1412–1456 | 847 | |
| 50 | #50Zglm-5.3-flash | Z.ai | 1433 | 1402–1464 | 408 |
Data: LMArena leaderboard dataset (CC BY 4.0). Changes: Top 50 entries per leaderboard; scores are rounded. https://huggingface.co/datasets/lmarena-ai/leaderboard-datasetCompany logos are trademarks of their respective owners and are used only for identification (icons: Simple Icons).Data: Epoch AI — Capabilities & benchmarking (CC BY 4.0). https://epoch.ai/benchmarksSource: OpenRouter (openrouter.ai/rankings), as of 2026-10-04. Licensed under CC BY 4.0. https://openrouter.ai/rankings