Valumigo

LMArena · Coding

We regularly fetch and display public benchmark data: LMArena user voting rankings, Epoch AI test scores, and OpenRouter usage rankings. Each leaderboard shows both its publication date and the date we retrieved it. These scores are not assigned by this site.

LMArena scores come from people comparing two models' answers side by side and voting for the better one, using an Elo-based system. If score differences fall within the confidence intervals, treat the models as roughly comparable.

How to read this leaderboard

What it measures: A ranking calculated from votes classified as coding-related questions in LMArena text conversations.

How to read the score: A relative score reflecting which models users preferred for coding answers. Interpret small score differences alongside confidence intervals and vote counts.

Caveats: This evaluates preferences in coding conversations, so distinguish it from agent evaluations that actually modify and test repositories. Also check SWE-bench, Terminal-Bench and published hands-on results.

Insights from this ranking

  • The 95% confidence intervals for 1st-place gemini-4-argon-high and 2nd-place claude-fable-5-high overlap. This aggregation alone does not clearly establish their order.
  • 14 other models have confidence intervals that overlap with the 1st-place model's. Interpret small ranking differences cautiously alongside vote counts.
  • Anthropic has the most models among the top 10, with 5.
  • The 1st-place model received 1,230 votes, and this leaderboard includes 408 models.

Coding LMArena

Published 2026-10-02 · Retrieved 2026-10-04

15001520154015601580
Googlegemini-4-argon-high
1560
Anthropicclaude-fable-5-high
1552
Anthropicclaude-opus-4-6-high
1551
Anthropicclaude-opus-4-7-high
1551
Ogpt-6-astra-max
1543
Ogpt-6.1-sol-max
1542
Moonshot AIkimi-k3-max
1541
Xiaomimimo-v2.6-pro
1540
Metamuse-spark-1.3-max
1539
Anthropicclaude-opus-5.5-high
1538
Anthropicclaude-sonnet-5.5-xhigh
1536
Ogpt-5.6-sol-xhigh
1534
Anthropicclaude-opus-5-high
1533
Anthropicclaude-opus-4-8-high
1533
Metamuse-spark-1.1
1533
Metamuse-spark-1.2 (xHigh)
1531
Anthropicclaude-fable-5.1-max
1531
Anthropicclaude-opus-4-5-20251101-high-32k
1530
Googlegemini-3.8-flash-high
1530
Metamuse-spark
1530

Dots show scores; horizontal lines show 95% confidence intervals. Overlapping intervals suggest similar performance. The chart shows only the highest-scoring reasoning setting, such as high or max, for each model. The table shows rankings for all settings. Chat and vision rankings use 'style and length adjustment (style control)', the same default as the LMArena website.

Top ranking by company

  • GoogleGooglegemini-4-argon-high#1
  • AnthropicAnthropicclaude-fable-5-high#2
  • OOpenAIgpt-6-astra-max#7
  • Moonshot AIMoonshot AIkimi-k3-max#9
  • XiaomiXiaomimimo-v2.6-pro#10
  • MetaMetamuse-spark-1.3-max#11
  • DeepSeekDeepSeekdeepseek-v4.1-flash-max#26
  • AlibabaAlibabaqwen3.7-max-preview#27
  • ZZ.aiglm-5.3-flash#29
View table (top 50)
RankModelDeveloperScore95% confidence intervalVotes
1#1Googlegemini-4-argon-highGoogle15601543–15771,230
2#2Anthropicclaude-fable-5-highAnthropic15521545–155910,038
3#3Anthropicclaude-opus-4-6-highAnthropic15511546–155720,235
4#4Anthropicclaude-opus-4-7-highAnthropic15511545–155618,521
5#5Anthropicclaude-opus-4-6Anthropic15481542–155322,901
6#6Anthropicclaude-opus-4-7Anthropic15461541–155218,728
7#7Ogpt-6-astra-maxOpenAI15431530–15562,135
8#8Ogpt-6.1-sol-maxOpenAI15421518–1566609
9#9Moonshot AIkimi-k3-maxMoonshot AI15411533–15497,285
10#10Xiaomimimo-v2.6-proXiaomi15401522–15581,103
11#11Metamuse-spark-1.3-maxMeta15391528–15493,307
12#12Anthropicclaude-opus-5.5-highAnthropic15381521–15561,060
13#13Anthropicclaude-sonnet-5.5-xhighAnthropic15361515–1558779
14#14Ogpt-5.6-sol-xhighOpenAI15341527–15419,943
15#15Anthropicclaude-opus-5-highAnthropic15331527–154015,559
16#16Anthropicclaude-opus-4-8-highAnthropic15331527–153917,395
17#17Metamuse-spark-1.1Meta15331526–153910,537
18#18Metamuse-spark-1.2 (xHigh)Meta15311514–15491,222
19#19Anthropicclaude-fable-5.1-maxAnthropic15311519–15442,358
20#20Anthropicclaude-opus-4-5-20251101-high-32kAnthropic15301523–15387,793
21#21Anthropicclaude-opus-5-maxAnthropic15301522–15387,382
22#22Googlegemini-3.8-flash-highGoogle15301522–15387,067
23#23Anthropicclaude-opus-4-8Anthropic15301524–153518,024
24#24Metamuse-sparkMeta15301520–15403,930
25#25Anthropicclaude-sonnet-4-6Anthropic15291523–153419,686
26#26DeepSeekdeepseek-v4.1-flash-maxDeepSeek15281516–15402,469
27#27Alibabaqwen3.7-max-previewAlibaba15241506–15421,165
28#28Alibabaqwen3.8-maxAlibaba15241516–15326,611
29#29Zglm-5.3-flashZ.ai15231515–15326,157
30#30Anthropicclaude-opus-4-5-20251101Anthropic15231518–152817,962
31#31Googlegemini-3.1-pro-previewGoogle15221517–152633,581
32#32Xiaomimimo-v2.5-proXiaomi15221516–152719,324
33#33Xiaomimimo-v2.6-flashXiaomi15211506–15371,508
34#34Zglm-5.3-maxZ.ai15211511–15304,424
35#35Ogpt-5.4-highOpenAI15211515–152717,566
36#36Ogpt-6-sol-maxOpenAI15201507–15342,071
37#37Anthropicclaude-sonnet-4-5-20250929-high-32kAnthropic15191514–152419,905
38#38Googlegemini-3.6-flash-highGoogle15191512–152610,350
39#39Anthropicclaude-sonnet-5-highAnthropic15191512–152512,377
40#40Ogpt-5.5-highOpenAI15191513–152419,110
41#41Googlegemini-3.7-flash-highGoogle15181510–15266,041
42#42Googlegemini-3-proGoogle15181511–15258,787
43#43Moonshot AIkimi-k2.6Moonshot AI15161509–152311,062
44#44Ogpt-5.6-terra-xhighOpenAI15151508–152210,166
45#45Ogpt-5.2-chat-latest-20260210OpenAI15151508–15229,629
46#46xgrok-4.5xAI15141507–152111,055
47#47ByteDancedola-seed-2.0-proByteDance15141509–151921,996
48#48Alibabaqwen3.5-max-previewAlibaba15141506–15226,358
49#49Ogpt-5.5-instantOpenAI15131506–15217,909
50#50Zglm-5.1Z.ai15131507–151816,060

Data: LMArena leaderboard dataset (CC BY 4.0). Changes: Top 50 entries per leaderboard; scores are rounded. https://huggingface.co/datasets/lmarena-ai/leaderboard-datasetCompany logos are trademarks of their respective owners and are used only for identification (icons: Simple Icons).Data: Epoch AI — Capabilities & benchmarking (CC BY 4.0). https://epoch.ai/benchmarksSource: OpenRouter (openrouter.ai/rankings), as of 2026-10-04. Licensed under CC BY 4.0. https://openrouter.ai/rankings