Valumigo

LMArena · Chinese

We regularly fetch and display public benchmark data: LMArena user voting rankings, Epoch AI test scores, and OpenRouter usage rankings. Each leaderboard shows both its publication date and the date we retrieved it. These scores are not assigned by this site.

LMArena scores come from people comparing two models' answers side by side and voting for the better one, using an Elo-based system. If score differences fall within the confidence intervals, treat the models as roughly comparable.

How to read this leaderboard

What it measures: A ranking calculated from votes on questions classified as Chinese in LMArena text conversations.

How to read the score: Shows which models voters preferred for questions in Chinese. The question's language does not indicate the voter's native language or identity.

Caveats: Vote counts and confidence interval widths vary by language and model. For new models, check vote counts, update dates and aggregation conditions together.

Insights from this ranking

  • The 95% confidence intervals for 1st-place gemini-4-argon-high and 2nd-place claude-opus-5.5-high overlap. This aggregation alone does not clearly establish their order.
  • 10 other models have confidence intervals that overlap with the 1st-place model's. Interpret small ranking differences cautiously alongside vote counts.
  • Anthropic has the most models among the top 10, with 7.
  • The 1st-place model received 352 votes, and this leaderboard includes 389 models.

Chinese LMArena

Published 2026-10-02 · Retrieved 2026-10-04

14801520156016001640
Googlegemini-4-argon-high
1587
Anthropicclaude-opus-5.5-high
1571
Anthropicclaude-fable-5.1-max
1570
Anthropicclaude-fable-5-high
1555
Anthropicclaude-opus-4-6-high
1551
Anthropicclaude-opus-5-max
1548
Googlegemini-3.7-flash-high
1546
Moonshot AIkimi-k3-max
1540
Alibabaqwen3.8-max
1538
Ogpt-5.6-sol-xhigh
1538
Anthropicclaude-opus-4-7-high
1537
Googlegemini-3.8-flash-high
1536
Anthropicclaude-sonnet-5.5-xhigh
1533
Metamuse-spark-1.1
1531
Metamuse-spark-1.3-max
1531
Zglm-5.3-max
1531
Googlegemini-3.6-flash-high
1530
Zglm-5.3-flash
1529
Googlegemini-3.1-pro-preview
1529
Metamuse-spark-1.2 (xHigh)
1528

Dots show scores; horizontal lines show 95% confidence intervals. Overlapping intervals suggest similar performance. The chart shows only the highest-scoring reasoning setting, such as high or max, for each model. The table shows rankings for all settings. Chat and vision rankings use 'style and length adjustment (style control)', the same default as the LMArena website.

Top ranking by company

  • GoogleGooglegemini-4-argon-high#1
  • AnthropicAnthropicclaude-opus-5.5-high#2
  • Moonshot AIMoonshot AIkimi-k3-max#10
  • AlibabaAlibabaqwen3.8-max#11
  • OOpenAIgpt-5.6-sol-xhigh#12
  • MetaMetamuse-spark-1.1#17
  • ZZ.aiglm-5.3-max#19
  • DeepSeekDeepSeekdeepseek-v4.1-flash-max#29
  • XiaomiXiaomimimo-v2.5-pro#39
View table (top 50)
RankModelDeveloperScore95% confidence intervalVotes
1#1Googlegemini-4-argon-highGoogle15871555–1619352
2#2Anthropicclaude-opus-5.5-highAnthropic15711542–1601402
3#3Anthropicclaude-fable-5.1-maxAnthropic15701549–1591874
4#4Anthropicclaude-fable-5-highAnthropic15551544–15672,852
5#5Anthropicclaude-opus-4-6-highAnthropic15511542–15605,342
6#6Anthropicclaude-opus-5-maxAnthropic15481535–15622,274
7#7Anthropicclaude-opus-5-highAnthropic15481538–15584,718
8#8Googlegemini-3.7-flash-highGoogle15461532–15611,681
9#9Anthropicclaude-opus-4-6Anthropic15441535–15535,582
10#10Moonshot AIkimi-k3-maxMoonshot AI15401527–15542,051
11#11Alibabaqwen3.8-maxAlibaba15381523–15531,712
12#12Ogpt-5.6-sol-xhighOpenAI15381526–15502,823
13#13Anthropicclaude-opus-4-7-highAnthropic15371528–15474,149
14#14Googlegemini-3.8-flash-highGoogle15361522–15502,061
15#15Anthropicclaude-opus-4-7Anthropic15341524–15444,133
16#16Anthropicclaude-sonnet-5.5-xhighAnthropic15331497–1568244
17#17Metamuse-spark-1.1Meta15311519–15422,943
18#18Metamuse-spark-1.3-maxMeta15311511–1550939
19#19Zglm-5.3-maxZ.ai15311514–15471,328
20#20Googlegemini-3.6-flash-highGoogle15301518–15422,712
21#21Zglm-5.3-flashZ.ai15291513–15441,648
22#22Googlegemini-3.1-pro-previewGoogle15291521–15368,287
23#23Metamuse-spark-1.2 (xHigh)Meta15281497–1559382
24#24Alibabaqwen3.7-max-previewAlibaba15271491–1562309
25#25Ogpt-5.5OpenAI15261517–15364,451
26#26Anthropicclaude-opus-4-8Anthropic15261516–15364,394
27#27Anthropicclaude-opus-4-8-highAnthropic15261516–15354,356
28#28Metamuse-sparkMeta15251505–1545965
29#29DeepSeekdeepseek-v4.1-flash-maxDeepSeek15231500–1546672
30#30Zglm-5.2-maxZ.ai15201510–15313,261
31#31Googlegemini-3-proGoogle15201510–15313,444
32#32Googlegemini-3.5-flash-mediumGoogle15171507–15283,451
33#33Ogpt-5.5-highOpenAI15171507–15274,469
34#34Alibabaqwen3.5-max-previewAlibaba15161501–15321,575
35#35Ogpt-5.6-terra-xhighOpenAI15161505–15282,949
36#36Googlegemini-3-flashGoogle15161504–15292,466
37#37Zglm-5Z.ai15161503–15302,001
38#38Moonshot AIkimi-k2.6Moonshot AI15161503–15292,326
39#39Xiaomimimo-v2.5-proXiaomi15151505–15244,455
40#40Googlegemini-3.5-flash-highGoogle15151504–15253,640
41#41Zglm-5.1Z.ai15141504–15244,041
42#42Xiaomimimo-v2.6-proXiaomi15111478–1543297
43#43Ogpt-6-astra-maxOpenAI15101488–1532674
44#44xgrok-4.20-beta1xAI15101495–15251,726
45#45Thy3Tencent15091489–1530779
46#46Anthropicclaude-sonnet-4-6Anthropic15091500–15194,501
47#47xgrok-4.7-xhighxAI15071482–1532531
48#48Ogpt-5.4-highOpenAI15071497–15174,132
49#49Ogpt-5.5-instantOpenAI15061490–15211,693
50#50xgrok-4.5xAI15041493–15163,061

Data: LMArena leaderboard dataset (CC BY 4.0). Changes: Top 50 entries per leaderboard; scores are rounded. https://huggingface.co/datasets/lmarena-ai/leaderboard-datasetCompany logos are trademarks of their respective owners and are used only for identification (icons: Simple Icons).Data: Epoch AI — Capabilities & benchmarking (CC BY 4.0). https://epoch.ai/benchmarksSource: OpenRouter (openrouter.ai/rankings), as of 2026-10-04. Licensed under CC BY 4.0. https://openrouter.ai/rankings