Valumigo

LMArena · 西班牙语

定期获取并展示公开的基准测试资料(LMArena 用户投票排名、Epoch AI 测试分数、OpenRouter 使用量排名)。每个榜单均标注数据发布日期和获取日期。这些分数并非本站自行评定。

LMArena 分数来自用户并排查看两个模型的回答,再投票选出更好的一方(Elo 方式)。如果分数差异在置信区间内,可视为水平基本相近。

如何解读此排行榜

衡量什么: 根据 LMArena 文本对话中被归类为 西班牙语 的问题所获投票计算的排名。

如何解读分数: 显示投票者在 西班牙语 问题中更偏好的模型。问题的语言并不代表投票者的母语或身份。

注意事项: 投票数量和置信区间宽度因语言和模型而异。对于新模型,请同时查看投票数量、更新日期及汇总条件。

本次排名的解读要点

  • 第 1 名 claude-fable-5-high 与第 2 名 claude-opus-4-6 的 95% 置信区间重叠。仅凭本次统计难以确定两个模型的先后顺序。
  • 还有 27 个模型的置信区间与第 1 名重叠。解读细微排名差异时,请结合投票数量保持谨慎。
  • 前 10 个模型中,Anthropic 的模型有 6 个,数量最多。
  • 第 1 名模型获得了 1,167 票,此排行榜包含 295 个模型。

西班牙语 LMArena

发布日期:2026-10-02 · 获取日期:2026-10-04

14401470150015301560
Anthropicclaude-fable-5-high
1507
Anthropicclaude-opus-4-6
1500
Anthropicclaude-opus-4-7-high
1499
Googlegemini-4-argon-high
1497
Googlegemini-3.7-flash-high
1494
Anthropicclaude-fable-5.1-max
1487
Ogpt-5.5-instant
1485
Moonshot AIkimi-k3-max
1485
Anthropicclaude-opus-5-max
1484
Alibabaqwen3.8-max
1483
Metamuse-spark-1.3-max
1482
Googlegemini-3.1-pro-preview
1481
Anthropicclaude-opus-4-8-high
1481
Metamuse-spark
1480
Googlegemini-3-flash
1478
DeepSeekdeepseek-v4.1-flash-max
1478
Zglm-5.2-max
1477
Googlegemini-3.8-flash-high
1476
Metamuse-spark-1.1
1476
Anthropicclaude-opus-4-5-20251101-high-32k
1476

圆点表示得分,横线表示 95% 置信区间。横线重叠意味着水平基本相近。 图表仅显示同一模型各推理设置(high、max 等)中排名最高的一项。表格列出所有设置各自的排名。 对话与视觉排名采用与 LMArena 网站默认设置相同的“语气与长度校正(style control)”排名。

各公司最高排名

  • AnthropicAnthropicclaude-fable-5-high#1
  • GoogleGooglegemini-4-argon-high#4
  • OOpenAIgpt-5.5-instant#8
  • Moonshot AIMoonshot AIkimi-k3-max#9
  • AlibabaAlibabaqwen3.8-max#11
  • MetaMetamuse-spark-1.3-max#12
  • DeepSeekDeepSeekdeepseek-v4.1-flash-max#17
  • ZZ.aiglm-5.2-max#18
  • BaiduBaiduernie-5.0-preview-1203#28
查看表格(前 50 名)
排名模型开发方分数95% 置信区间投票数
1#1Anthropicclaude-fable-5-highAnthropic15071489–15251,167
2#2Anthropicclaude-opus-4-6Anthropic15001487–15132,707
3#3Anthropicclaude-opus-4-7-highAnthropic14991484–15132,043
4#4Googlegemini-4-argon-highGoogle14971456–1538201
5#5Googlegemini-3.7-flash-highGoogle14941471–1518639
6#6Anthropicclaude-opus-4-6-highAnthropic14931480–15062,627
7#7Anthropicclaude-fable-5.1-maxAnthropic14871457–1516375
8#8Ogpt-5.5-instantOpenAI14851463–1507811
9#9Moonshot AIkimi-k3-maxMoonshot AI14851465–1505843
10#10Anthropicclaude-opus-5-maxAnthropic14841463–1506803
11#11Alibabaqwen3.8-maxAlibaba14831459–1506689
12#12Metamuse-spark-1.3-maxMeta14821451–1514365
13#13Googlegemini-3.1-pro-previewGoogle14811470–14923,868
14#14Anthropicclaude-opus-4-8-highAnthropic14811466–14961,840
15#15Metamuse-sparkMeta14801453–1507518
16#16Googlegemini-3-flashGoogle14781460–1497999
17#17DeepSeekdeepseek-v4.1-flash-maxDeepSeek14781442–1513274
18#18Zglm-5.2-maxZ.ai14771460–14941,314
19#19Googlegemini-3.8-flash-highGoogle14761456–1497848
20#20Metamuse-spark-1.1Meta14761458–14951,106
21#21Anthropicclaude-opus-4-5-20251101-high-32kAnthropic14761459–14941,137
22#22Googlegemini-3.6-flash-highGoogle14761456–1497914
23#23Ogpt-5.2-chat-latest-20260210OpenAI14751457–14921,206
24#24Anthropicclaude-opus-4-7Anthropic14741460–14882,099
25#25Anthropicclaude-opus-5-highAnthropic14741458–14901,676
26#26Googlegemini-3-proGoogle14721454–14891,215
27#27Googlegemini-3.5-flash-highGoogle14711455–14881,410
28#28Baiduernie-5.0-preview-1203Baidu14701433–1507238
29#29Ogpt-5.5-highOpenAI14691455–14841,937
30#30ByteDancedola-seed-2.0-proByteDance14691456–14822,565
31#31Thy3Tencent14691439–1498371
32#32xgrok-4.20-beta1xAI14681449–1488952
33#33Anthropicclaude-opus-4-1-20250805Anthropic14681454–14812,293
34#34Ogpt-5.6-sol-xhighOpenAI14671448–14861,040
35#35Googlegemini-3.5-flash-mediumGoogle14671450–14851,274
36#36Anthropicclaude-sonnet-4-6Anthropic14661452–14802,177
37#37Anthropicclaude-opus-4-8Anthropic14661452–14811,918
38#38Anthropicclaude-opus-4-1-20250805-thinking-16kAnthropic14661448–14831,292
39#39Zglm-5.3-flashZ.ai14661444–1487792
40#40xgrok-4.1-thinkingxAI14661451–14801,944
41#41xgrok-4.20-multi-agent-beta-0309xAI14651451–14801,997
42#42Anthropicclaude-opus-4-5-20251101Anthropic14651452–14792,156
43#43Anthropicclaude-sonnet-4-5-20250929Anthropic14651453–14782,487
44#44Zglm-5.1Z.ai14651450–14791,897
45#45Xiaomimimo-v2.5-proXiaomi14641450–14792,126
46#46Ogpt-5.6-luna-xhighOpenAI14631444–14821,070
47#47xgrok-4.20-beta-0309-reasoningxAI14621448–14762,057
48#48Ogpt-5.6-terra-xhighOpenAI14621443–14811,083
49#49Ogpt-5.4OpenAI14611446–14752,037
50#50Moonshot AIkimi-k2.6Moonshot AI14601443–14761,371

数据:LMArena leaderboard dataset(CC BY 4.0)。修改:各榜单仅显示前 50 名,分数四舍五入。 https://huggingface.co/datasets/lmarena-ai/leaderboard-dataset公司标志为各公司的商标,仅用于帮助区分(图标:Simple Icons)。数据:Epoch AI — Capabilities & benchmarking(CC BY 4.0)。 https://epoch.ai/benchmarksSource: OpenRouter (openrouter.ai/rankings), as of 2026-10-04. Licensed under CC BY 4.0. https://openrouter.ai/rankings