Grok 4.6
xAI · 2026-08-12
Basic Information
- Developer
- xxAI
- Release Date
- 2026-08-12
- Overall Capability Index (ECI)
- 156.6
Benchmark Scores (Epoch AI)
Each benchmark also shows the model's rank among the models evaluated.
xPhD-level science questions (GPQA Diamond) · #10 of 30
94.0%
xAdvanced mathematics (FrontierMath) · #27 of 30
66.0%
xReasoning through unfamiliar puzzles (ARC-AGI-2) · #20 of 30
67.1%
xFactual question accuracy (SimpleQA Verified) · #27 of 30
49.3%
Bars show accuracy (%) and start at 0%.
User Vote Rankings (LMArena)
Rankings based on people's votes comparing responses from two models. Models do not appear on leaderboards where they have too few votes.
- Overall#73 of 413 · 1454
- English#82 of 413 · 1458
- Web development#23 of 50 · 1620
- Document understanding#25 of 44 · 1452
- Overall agent performance#25 of 50 · 0.013
- Agent task success#31 of 50 · -0.006
- French#36 of 290 · 1491
- Korean#41 of 278 · 1431
- Image understanding (vision)#41 of 50 · 1263
- Creative writing#52 of 411 · 1445
- Chinese#53 of 389 · 1503
- German#59 of 307 · 1457
- Coding#64 of 408 · 1507
- Math#78 of 396 · 1448
- Japanese#82 of 274 · 1402
- Spanish#102 of 295 · 1429
Where to Use It
This site has not yet compiled the services and plans that offer this model. Check the developer's official information.
Scores and rankings are reproduced directly from public sources and were not measured by this site. Model names vary slightly across sources and are matched automatically, so different versions may occasionally be mixed together.
Source: Epoch AI (CC BY 4.0) · Checked 2026-10-04 · LMArena (CC BY 4.0) · Checked 2026-10-04