Skip to content

每周AI性能报告

本周 September 23, 2026

50 个模型上升, 245 个下降, 5 个不变,本周共追踪 300 个模型。

1

执行摘要

最大赢家

Claude Opus 5.5

+430 个名次

当前排名 #5

最大输家

Ternary Bonsai 2 27B

-24 个名次

当前排名 #250

新上榜

10

个新模型本周进入排名

2

每周变动

过去7天内排名绝对变动最大的10个模型。

模型评分7天变化排名
Claude Opus 5.5Anthropic95.1+430#5
Grok 4.7xAI88.8+392#43
MiMo-V2.6-ProXiaomi76.2+318#117
Claude Opus 5Anthropic95.1+272#6
Claude Sonnet 5Anthropic85.2+231#63
gpt-oss-20b (batch)OpenAI57.4+215#220
Aion 3.5 Miniaion-labs40.0+200#235
Aion 3.5aion-labs40.0+199#236
Solar Mini 4Upstage40.0+198#237
GPT-6 Luna ProOpenAI40.0+197#238
3

涨幅最大

本周排名提升最大的模型。

1+430
2+392
3+318
4
Claude Opus 5Anthropic
+272
5+231
7
Aion 3.5 Miniaion-labs
+200
8
Aion 3.5aion-labs
+199
9+198
10+197
4

跌幅最大

本周排名下降最大的模型。

2
Paretounbiased
-24
3-24
5
Schematron V2 Turboinference-net
-24
6
Schematron V2 Smallinference-net
-24
7-24
8-24
9-24
10-24
5

新模型

10 个新模型本周进入排名。

模型评分排名
Claude Opus 5.5Anthropic95.1#5
Grok 4.7xAI88.8#43
MiMo-V2.6-ProXiaomi76.2#117
gpt-oss-20b (batch)OpenAI57.4#220
Aion 3.5 Miniaion-labs40.0#235
Aion 3.5aion-labs40.0#236
Solar Mini 4Upstage40.0#237
GPT-6 Luna ProOpenAI40.0#238
GPT-6 Luna Pro (batch)OpenAI40.0#239
GPT-6 LunaOpenAI40.0#240
6

关注列表

处于脆弱状态可能进一步恶化的模型。需密切关注。

模型评分7天变化
95.1+272
92.7+8
90.5+18
90.5+80
88.8+56
88.8+30
88.8-8
88.7-10
88.7-10
88.7-10
88.7-10
88.4-10
88.4-10
87.9-9
87.8-7
87.8-6
86.7-6
86.7-6
86.7-6
85.2+231
85.2-7
85.2-7
85.1-7
85.1-7
83.5-7
83.5-7
83.5-7
83.4-7
82.4-7
82.4-7
81.8-6
81.8-6
81.8-6
81.8-6
81.4-6
81.4-6
81.0-6
81.0-6
80.9+161
80.9+189
80.9-7
80.5-7
80.5-7
79.4-7
79.4-7
79.3-7
79.3-7
79.3-7
79.3-7
79.3-7
79.1-7
79.1-7
79.1-7
79.1-7
79.0-7
79.0-7
78.6-8
78.6-8
78.0-7
78.0+11
78.0+43
78.0-9
77.9-9
77.9-9
77.7-8
77.4-8
77.0-6
76.8-6
76.8-6
76.5-6
76.5-6
76.2-7
76.0-7
75.7-7
75.7-7
75.6+120
75.6-8
75.3-7
75.2-7
75.1+13
75.1+27
75.1-9
74.8+176
74.8+187
74.8-11
74.4-11
74.4-11
74.4-10
74.3-10
74.1-10
73.9-10
73.9-15
73.8-11
73.6-10
73.4-10
73.0-10
73.0-10
73.0-10
72.9-10
72.9-10
72.9-10
72.3-8
72.1-8
72.0-8
71.8+7
71.8-10
71.6-10
71.6+6
71.6-11
71.2+60
71.2-12
71.2-12
71.2-12
71.0-12
70.8-12
70.7-12
70.7-12
69.9-9
69.9-9
69.5-8
69.3-8
69.3-8
69.3-8
68.7-8
68.7-8
68.6-8
68.4-8
68.2+159
68.2-9
68.1-9
67.7-9
67.7-9
67.4-8
67.0-8
66.8-8
66.7+6
66.7-9
66.7-9
66.5-9
66.1-9
65.9+23
65.9-10
65.7-10
65.7-10
65.6-10
65.3-10
65.3-10
64.9-10
64.7-10
64.1-9
64.1+170
64.1-12
63.9-10
63.5-10
63.4-10
63.1-10
62.5-10
62.3-10
61.9-10
61.6-10
61.5-9
61.2-9
61.0-9
60.9-9
60.4-9
60.2-9
60.2-9
58.9-9
58.9-9
57.4-9
56.5-10
56.2-10
54.0-9
53.8-9
53.8-9
53.2-7
51.3-9
50.8-9
48.7-8
48.7-9
47.5-9
44.5-9
42.1-9
40.7-9
40.0-24
40.0-24
40.0-24
40.0-24
40.0-24
40.0-24
40.0-24
40.0-24
40.0-24
40.0-24
40.0-24
40.0-24
40.0-24
40.0-23
40.0-24
40.0-23
40.0-22
40.0-22
40.0-22
40.0-22
40.0-22
40.0-22
40.0-22
40.0-21
40.0-21
40.0-21
40.0-21
40.0-21
40.0-21
40.0-21
40.0-21
40.0-21
40.0-20
40.0-20
40.0-19
40.0-19
40.0-19
40.0-19
40.0-19
40.0-19
40.0-17
40.0-17
40.0-16
40.0-15
40.0-15
40.0-15
40.0-15
40.0-15
40.0-15
40.0-15
7

排行榜快照

本周综合评分前10的AI模型。

#模型评分7天变化
1Claude Fable 5.1Anthropic95.90
2Claude Fable 5.1 (batch)Anthropic95.90
3Claude Fable 5Anthropic95.90
4Claude Fable 5 (batch)Anthropic95.90
5Claude Opus 5.5Anthropic95.1+430
6Claude Opus 5Anthropic95.1+272
7Claude Opus 4.8Anthropic95.10
8Claude Opus 4.7Anthropic95.1-3
9Claude Opus 4.7 (batch)Anthropic95.1-3
10Claude Opus 4.8 (batch)Anthropic94.6-2
8

本周数据

总排名数

300

平均评分

68.2

前10变动

5/10

最活跃的服务商

OpenAI

9

阅读更多分析

Frequently Asked Questions

The weekly report summarizes seven days of AI model ranking changes. It highlights the biggest gainers and losers by rank position, new models entering the leaderboard, models in a fragile watch-list state, and a snapshot of the current top 10. All data comes from hourly-updated composite scores across 290+ models.

Gainers and losers are ranked by their 7-day rank change, which measures how many positions a model moved up or down on the leaderboard over the past week. The models with the largest positive changes are the top gainers, and those with the largest negative changes are the top losers.

The watch list contains models currently in a "fragile" state, meaning their rankings are unstable and could degrade further. Monitoring these models is important if you rely on them in production, as they may experience significant quality or performance drops in the near term.

Weekly AI Performance Report - This Week in AI Models | LM Market Cap