下拉刷新
August 2026 LM Ranking
rank_logo
RankModelScore
1claude-fable-5
1506
2claude-opus-4-6-thinking
1505
3claude-opus-4-7-thinking
1502
4muse-spark-1.2 (xHigh)
1498
5claude-opus-4-6
1498
6claude-opus-5-high
1494
7claude-opus-4-7
1494
8claude-opus-5-max
1490
9qwen3.8-max
1490
10muse-spark-1.1
1489
11muse-spark
1488
12kimi-k3-max
1487
13gemini-3.1-pro
1486
14gemini-3-pro
1485
15gemini-3.6-flash
1484
16gpt-5.5-high
1482
17claude-opus-4-8-thinking
1481
18gpt-5.6-sol-xhigh
1481
19gemini-3.5-flash-high
1477
20gpt-5.5
1477

The LMArena Ranking is a crowdsourced leaderboard for large language models. Users chat with two anonymous models and vote for the better response, with model ratings calculated using the Elo rating system. The leaderboard covers multiple capability dimensions including text, vision, and code, making it one of the most authoritative LLM evaluation benchmarks. Based on this ranking, we have done model name aggregation and cleaning work.