Valumigo

LMArena · Spanish

We regularly fetch and display public benchmark data: LMArena user voting rankings, Epoch AI test scores, and OpenRouter usage rankings. Each leaderboard shows both its publication date and the date we retrieved it. These scores are not assigned by this site.

LMArena scores come from people comparing two models' answers side by side and voting for the better one, using an Elo-based system. If score differences fall within the confidence intervals, treat the models as roughly comparable.

How to read this leaderboard

What it measures: A ranking calculated from votes on questions classified as Spanish in LMArena text conversations.

How to read the score: Shows which models voters preferred for questions in Spanish. The question's language does not indicate the voter's native language or identity.

Caveats: Vote counts and confidence interval widths vary by language and model. For new models, check vote counts, update dates and aggregation conditions together.

Insights from this ranking

  • The 95% confidence intervals for 1st-place claude-fable-5-high and 2nd-place claude-opus-4-6 overlap. This aggregation alone does not clearly establish their order.
  • 27 other models have confidence intervals that overlap with the 1st-place model's. Interpret small ranking differences cautiously alongside vote counts.
  • Anthropic has the most models among the top 10, with 6.
  • The 1st-place model received 1,167 votes, and this leaderboard includes 295 models.

Spanish LMArena

Published 2026-10-02 · Retrieved 2026-10-04

14401470150015301560
Anthropicclaude-fable-5-high
1507
Anthropicclaude-opus-4-6
1500
Anthropicclaude-opus-4-7-high
1499
Googlegemini-4-argon-high
1497
Googlegemini-3.7-flash-high
1494
Anthropicclaude-fable-5.1-max
1487
Ogpt-5.5-instant
1485
Moonshot AIkimi-k3-max
1485
Anthropicclaude-opus-5-max
1484
Alibabaqwen3.8-max
1483
Metamuse-spark-1.3-max
1482
Googlegemini-3.1-pro-preview
1481
Anthropicclaude-opus-4-8-high
1481
Metamuse-spark
1480
Googlegemini-3-flash
1478
DeepSeekdeepseek-v4.1-flash-max
1478
Zglm-5.2-max
1477
Googlegemini-3.8-flash-high
1476
Metamuse-spark-1.1
1476
Anthropicclaude-opus-4-5-20251101-high-32k
1476

Dots show scores; horizontal lines show 95% confidence intervals. Overlapping intervals suggest similar performance. The chart shows only the highest-scoring reasoning setting, such as high or max, for each model. The table shows rankings for all settings. Chat and vision rankings use 'style and length adjustment (style control)', the same default as the LMArena website.

Top ranking by company

  • AnthropicAnthropicclaude-fable-5-high#1
  • GoogleGooglegemini-4-argon-high#4
  • OOpenAIgpt-5.5-instant#8
  • Moonshot AIMoonshot AIkimi-k3-max#9
  • AlibabaAlibabaqwen3.8-max#11
  • MetaMetamuse-spark-1.3-max#12
  • DeepSeekDeepSeekdeepseek-v4.1-flash-max#17
  • ZZ.aiglm-5.2-max#18
  • BaiduBaiduernie-5.0-preview-1203#28
View table (top 50)
RankModelDeveloperScore95% confidence intervalVotes
1#1Anthropicclaude-fable-5-highAnthropic15071489–15251,167
2#2Anthropicclaude-opus-4-6Anthropic15001487–15132,707
3#3Anthropicclaude-opus-4-7-highAnthropic14991484–15132,043
4#4Googlegemini-4-argon-highGoogle14971456–1538201
5#5Googlegemini-3.7-flash-highGoogle14941471–1518639
6#6Anthropicclaude-opus-4-6-highAnthropic14931480–15062,627
7#7Anthropicclaude-fable-5.1-maxAnthropic14871457–1516375
8#8Ogpt-5.5-instantOpenAI14851463–1507811
9#9Moonshot AIkimi-k3-maxMoonshot AI14851465–1505843
10#10Anthropicclaude-opus-5-maxAnthropic14841463–1506803
11#11Alibabaqwen3.8-maxAlibaba14831459–1506689
12#12Metamuse-spark-1.3-maxMeta14821451–1514365
13#13Googlegemini-3.1-pro-previewGoogle14811470–14923,868
14#14Anthropicclaude-opus-4-8-highAnthropic14811466–14961,840
15#15Metamuse-sparkMeta14801453–1507518
16#16Googlegemini-3-flashGoogle14781460–1497999
17#17DeepSeekdeepseek-v4.1-flash-maxDeepSeek14781442–1513274
18#18Zglm-5.2-maxZ.ai14771460–14941,314
19#19Googlegemini-3.8-flash-highGoogle14761456–1497848
20#20Metamuse-spark-1.1Meta14761458–14951,106
21#21Anthropicclaude-opus-4-5-20251101-high-32kAnthropic14761459–14941,137
22#22Googlegemini-3.6-flash-highGoogle14761456–1497914
23#23Ogpt-5.2-chat-latest-20260210OpenAI14751457–14921,206
24#24Anthropicclaude-opus-4-7Anthropic14741460–14882,099
25#25Anthropicclaude-opus-5-highAnthropic14741458–14901,676
26#26Googlegemini-3-proGoogle14721454–14891,215
27#27Googlegemini-3.5-flash-highGoogle14711455–14881,410
28#28Baiduernie-5.0-preview-1203Baidu14701433–1507238
29#29Ogpt-5.5-highOpenAI14691455–14841,937
30#30ByteDancedola-seed-2.0-proByteDance14691456–14822,565
31#31Thy3Tencent14691439–1498371
32#32xgrok-4.20-beta1xAI14681449–1488952
33#33Anthropicclaude-opus-4-1-20250805Anthropic14681454–14812,293
34#34Ogpt-5.6-sol-xhighOpenAI14671448–14861,040
35#35Googlegemini-3.5-flash-mediumGoogle14671450–14851,274
36#36Anthropicclaude-sonnet-4-6Anthropic14661452–14802,177
37#37Anthropicclaude-opus-4-8Anthropic14661452–14811,918
38#38Anthropicclaude-opus-4-1-20250805-thinking-16kAnthropic14661448–14831,292
39#39Zglm-5.3-flashZ.ai14661444–1487792
40#40xgrok-4.1-thinkingxAI14661451–14801,944
41#41xgrok-4.20-multi-agent-beta-0309xAI14651451–14801,997
42#42Anthropicclaude-opus-4-5-20251101Anthropic14651452–14792,156
43#43Anthropicclaude-sonnet-4-5-20250929Anthropic14651453–14782,487
44#44Zglm-5.1Z.ai14651450–14791,897
45#45Xiaomimimo-v2.5-proXiaomi14641450–14792,126
46#46Ogpt-5.6-luna-xhighOpenAI14631444–14821,070
47#47xgrok-4.20-beta-0309-reasoningxAI14621448–14762,057
48#48Ogpt-5.6-terra-xhighOpenAI14621443–14811,083
49#49Ogpt-5.4OpenAI14611446–14752,037
50#50Moonshot AIkimi-k2.6Moonshot AI14601443–14761,371

Data: LMArena leaderboard dataset (CC BY 4.0). Changes: Top 50 entries per leaderboard; scores are rounded. https://huggingface.co/datasets/lmarena-ai/leaderboard-datasetCompany logos are trademarks of their respective owners and are used only for identification (icons: Simple Icons).Data: Epoch AI — Capabilities & benchmarking (CC BY 4.0). https://epoch.ai/benchmarksSource: OpenRouter (openrouter.ai/rankings), as of 2026-10-04. Licensed under CC BY 4.0. https://openrouter.ai/rankings