Valumigo

LMArena · 画像理解(ビジョン)

公開されているベンチマーク資料(LMArenaのユーザー投票順位、Epoch AIのテストスコア、OpenRouterの使用量順位)を定期的に取得して表示します。各ランキングの公開日と取得日を併記します。このサイトが独自に採点したものではありません。

LMArenaのスコアは、2つのモデルの回答を並べて見たユーザーが、よりよい方に投票した結果(Elo方式)です。スコアの差が信頼区間内なら、実質的には同程度と考えましょう。

このランキングの読み方

測定する能力: 画像を含む質問に対する2つのモデルの回答を比較し、投票した画像理解ランキングです。

スコアの読み方: ユーザーがより好んだ回答の相対スコアです。写真・図表・スクリーンショットの読み取りの客観的な正答率ではありません。

注意点: 文書内の小さな文字やKorean (Hangul)の文字認識など、必要な条件での正確性は別途確認してください。

今回のランキングから読み取れること

  • 1位のclaude-fable-5-highと2位のqwen3.8-maxの95%信頼区間は重なっています。この集計だけで2つのモデルの順位を確定するのは困難です。
  • 1位と信頼区間が重なるモデルが、ほかに11件あります。小さな順位差は、投票数も踏まえて慎重に解釈してください。
  • 上位10モデルのうち、Anthropicのモデルが5モデルで最も多くなっています。
  • 1位のモデルが獲得した投票は13,709票で、このランキングには50件のモデルがあります。

画像理解(ビジョン) LMArena

公開日 2026-10-02 · 取得日 2026-10-04

1260128013001320
Anthropicclaude-fable-5-high
1309
Alibabaqwen3.8-max
1301
Anthropicclaude-opus-4-6-high
1299
Anthropicclaude-opus-4-7
1298
Googlegemini-3.7-flash-high
1296
Metamuse-spark
1294
Metamuse-spark-1.2 (xHigh)
1293
Ogpt-6.1-sol-max
1291
Googlegemini-3.8-flash-high
1290
Metamuse-spark-1.3-max
1290
Googlegemini-3-pro
1289
Anthropicclaude-opus-5-high
1289
Anthropicclaude-fable-5.1-max
1288
Ogpt-5.5
1287
Anthropicclaude-opus-4-8-high
1286
Ogpt-5.6-sol-xhigh
1285
Googlegemini-3.5-flash-medium
1284
Ogpt-6-astra-max
1284
Ogpt-5.4-high
1284
Metamuse-spark-1.1
1281

点はスコア、横線は95%信頼区間です。線が重なる場合、実質的に同程度です。 グラフには、同じモデルの推論設定(high・maxなど)のうち、スコアが最も高いものだけを表示しています。表には設定ごとの順位をすべて掲載しています。 会話・ビジョンのランキングは、LMArenaサイトのデフォルトと同じ「口調・長さ補正(style control)」ランキングです。

企業別の最高順位

  • AnthropicAnthropicclaude-fable-5-high#1
  • AlibabaAlibabaqwen3.8-max#2
  • GoogleGooglegemini-3.7-flash-high#6
  • MetaMetamuse-spark#8
  • OOpenAIgpt-6.1-sol-max#10
  • xxAIgrok-4.5#28
  • ZZ.aiglm-5.3-flash#30
  • Moonshot AIMoonshot AIkimi-k2.6#38
  • SStepFunStep 5 Preview#45
表で見る(上位50件)
順位モデル開発元スコア95%信頼区間投票数
1#1Anthropicclaude-fable-5-highAnthropic13091301–131613,709
2#2Alibabaqwen3.8-maxAlibaba13011294–130910,040
3#3Anthropicclaude-opus-4-6-highAnthropic12991293–130622,328
4#4Anthropicclaude-opus-4-7Anthropic12981292–130523,169
5#5Anthropicclaude-opus-4-7-highAnthropic12981291–130422,755
6#6Googlegemini-3.7-flash-highGoogle12961286–13063,814
7#7Anthropicclaude-opus-4-6Anthropic12941288–130126,869
8#8Metamuse-sparkMeta12941284–13035,868
9#9Metamuse-spark-1.2 (xHigh)Meta12931279–13071,951
10#10Ogpt-6.1-sol-maxOpenAI12911275–13071,408
11#11Googlegemini-3.8-flash-highGoogle12901279–13023,086
12#12Metamuse-spark-1.3-maxMeta12901280–13013,843
13#13Googlegemini-3-proGoogle12891282–129713,620
14#14Anthropicclaude-opus-5-highAnthropic12891282–129615,900
15#15Anthropicclaude-fable-5.1-maxAnthropic12881278–12984,312
16#16Ogpt-5.5OpenAI12871281–129227,223
17#17Anthropicclaude-opus-4-8-highAnthropic12861280–129318,350
18#18Ogpt-5.6-sol-xhighOpenAI12851277–129210,421
19#19Googlegemini-3.5-flash-mediumGoogle12841277–129112,467
20#20Ogpt-6-astra-maxOpenAI12841273–12963,175
21#21Googlegemini-3.5-flash-highGoogle12841277–129112,336
22#22Ogpt-5.4-highOpenAI12841278–129029,177
23#23Metamuse-spark-1.1Meta12811274–128811,255
24#24Ogpt-5.5-highOpenAI12811275–128725,488
25#25Googlegemini-3.6-flash-highGoogle12801271–12887,392
26#26Anthropicclaude-opus-4-8Anthropic12791273–128618,857
27#27Googlegemini-3.1-pro-previewGoogle12791274–128545,558
28#28xgrok-4.5xAI12791272–128611,039
29#29Ogpt-5.4OpenAI12791272–128522,828
30#30Zglm-5.3-flashZ.ai12781269–12885,479
31#31Ogpt-5.2-chat-latest-20260210OpenAI12781271–128516,480
32#32Ogpt-5.5-instantOpenAI12761267–12857,605
33#33Anthropicclaude-sonnet-4-6Anthropic12751269–128127,429
34#34Googlegemini-3-flashGoogle12731268–127839,142
35#35Anthropicclaude-sonnet-5.5-xhighAnthropic12681253–12841,608
36#36Ogpt-5.6-terra-xhighOpenAI12681261–127610,292
37#37Alibabaqwen3.7-plusAlibaba12651259–127213,065
38#38Moonshot AIkimi-k2.6Moonshot AI12651258–127216,387
39#39Anthropicclaude-sonnet-5-highAnthropic12641257–127113,029
40#40Googlegemini-3.5-flash-liteGoogle12641256–12737,579
41#41xgrok-4.6-highxAI12631254–12735,736
42#42Googlegemma-4-31bGoogle12621256–126839,111
43#43Googlegemini-3-flash (thinking-minimal)Google12601254–126537,184
44#44Ogpt-5.6-luna-xhighOpenAI12601252–126710,446
45#45SStep 5 PreviewStepFun12601235–1284604
46#46ByteDancedola-seed-2.0-proByteDance12571249–126410,981
47#47xgrok-4.20-beta-0309-reasoningxAI12551249–126226,832
48#48Moonshot AIkimi-k2.5-thinkingMoonshot AI12521246–125828,326
49#49Ogpt-5.1-highOpenAI12511244–12599,842
50#50Ogpt-5.4-mini-highOpenAI12511244–125725,734

データ: LMArena leaderboard dataset(CC BY 4.0)。変更: 各ランキングは上位50件のみ表示。スコアは四捨五入。 https://huggingface.co/datasets/lmarena-ai/leaderboard-dataset企業ロゴは各社の商標であり、識別のためにのみ使用しています(アイコン:Simple Icons)。データ:Epoch AI — Capabilities & benchmarking(CC BY 4.0)。 https://epoch.ai/benchmarksSource: OpenRouter (openrouter.ai/rankings), as of 2026-10-04. Licensed under CC BY 4.0. https://openrouter.ai/rankings