2026 年 9 月大模型排行榜

| 排名 | 模型 | 分数 |
|---|---|---|
| 1 | gemini-4-argon-high | 1525 |
| 2 | claude-opus-4-6-high | 1505 |
| 3 | claude-fable-5-high | 1505 |
| 4 | claude-opus-5.5-high | 1504 |
| 5 | claude-opus-4-7-high | 1502 |
| 6 | claude-fable-5.1-max | 1501 |
| 7 | claude-opus-4-6 | 1497 |
| 8 | muse-spark-1.3-max | 1495 |
| 9 | claude-opus-4-7 | 1494 |
| 10 | muse-spark-1.2 (xHigh) | 1494 |
| 11 | gemini-3.8-flash-high | 1494 |
| 12 | muse-spark-1.1 | 1492 |
| 13 | claude-opus-5-high | 1491 |
| 14 | claude-opus-5-max | 1489 |
| 15 | muse-spark | 1489 |
| 16 | gemini-3.7-flash-high | 1488 |
| 17 | kimi-k3-max | 1488 |
| 18 | gemini-3.1-pro | 1487 |
| 19 | gemini-3-pro | 1485 |
| 20 | gpt-5.6-sol-xhigh | 1484 |
| 排名 | 模型 | 分数 | 机构 |
|---|---|---|---|
| 1 | gemini-4-argon-high | 1525 | |
| 2 | claude-opus-4-6-high | 1505 | Anthropic |
| 3 | claude-fable-5-high | 1505 | Anthropic |
| 4 | claude-opus-5.5-high | 1504 | Anthropic |
| 5 | claude-opus-4-7-high | 1502 | Anthropic |
| 6 | claude-fable-5.1-max | 1501 | Anthropic |
| 7 | claude-opus-4-6 | 1497 | Anthropic |
| 8 | muse-spark-1.3-max | 1495 | Meta |
| 9 | claude-opus-4-7 | 1494 | Anthropic |
| 10 | muse-spark-1.2 (xHigh) | 1494 | Meta |
| 11 | gemini-3.8-flash-high | 1494 | |
| 12 | muse-spark-1.1 | 1492 | Meta |
| 13 | claude-opus-5-high | 1491 | Anthropic |
| 14 | claude-opus-5-max | 1489 | Anthropic |
| 15 | muse-spark | 1489 | Meta |
| 16 | gemini-3.7-flash-high | 1488 | |
| 17 | kimi-k3-max | 1488 | Moonshot |
| 18 | gemini-3.1-pro | 1487 | |
| 19 | gemini-3-pro | 1485 | |
| 20 | gpt-5.6-sol-xhigh | 1484 | OpenAI |
「LMArena 排名」是基于众包用户投票的大语言模型排行榜。通过让用户与两个匿名模型对话并选择更好的回答,使用 Elo 评分系统计算模型的相对实力。该排行榜覆盖文本、视觉、代码等多个能力维度,是目前最权威的 LLM 评测榜单之一,基于此榜单我们做了模型名称聚合和清理工作。