Skip to content

每周AI性能报告

本周 August 9, 2026

81 个模型上升, 216 个下降, 3 个不变,本周共追踪 300 个模型。

1

执行摘要

最大赢家

Claude Fable 5 (batch)

+377 个名次

当前排名 #2

最大输家

Seed 1.6 Flash

-65 个名次

当前排名 #297

新上榜

10

个新模型本周进入排名

2

每周变动

过去7天内排名绝对变动最大的10个模型。

模型评分7天变化排名
Claude Fable 5 (batch)Anthropic97.1+377#2
Claude Opus 4.7 (batch)Anthropic95.1+370#9
Claude Opus 4.8 (batch)Anthropic94.6+369#10
GPT-5.5 Pro (batch)OpenAI92.7+367#12
GPT-5.5 (batch)OpenAI92.7+365#14
Gemini 3.1 Pro Preview (batch)Google92.2+362#17
GPT-5.4 Pro (batch)OpenAI91.9+360#19
GPT-5.4 (batch)OpenAI91.9+358#21
GPT-5.2 Pro (batch)OpenAI90.5+352#27
GPT-5.2 (batch)OpenAI90.5+350#29
3

涨幅最大

本周排名提升最大的模型。

4

跌幅最大

本周排名下降最大的模型。

2-70
3-69
4-69
5-69
7
Mistral NemoMistral AI
-69
9-67
10-67
5

新模型

10 个新模型本周进入排名。

6

关注列表

处于脆弱状态可能进一步恶化的模型。需密切关注。

模型评分7天变化
95.1+171
95.1+171
92.7-7
92.2-8
92.2-8
91.9-9
91.9-10
90.5+10
90.5-12
90.5+40
90.5-13
90.5-14
90.4-15
89.0-16
89.0-17
89.0-18
89.0-19
89.0-20
89.0-21
88.7-19
88.7-19
88.7-19
88.7-26
88.7-28
88.4-29
87.8-28
86.7-27
86.7-27
86.7-28
85.2+123
85.2-30
85.1-31
83.5-32
83.5-33
83.5-33
82.4-33
82.1-35
81.4-36
81.3-37
80.5-36
80.5-36
80.2-40
80.0-38
79.4-39
79.4-39
79.3-39
79.3-40
79.3-41
79.1-41
79.1-42
79.0-43
78.5-44
78.0-32
78.0-8
78.0-48
78.0-48
77.7-46
77.4-46
77.0-45
76.9-45
76.2-47
76.0-47
75.9-50
75.8-48
75.3-47
75.1-33
75.1-21
75.1-49
74.8+71
74.8+83
74.8-51
74.4-50
74.4-50
74.3-51
74.2-52
74.1-55
74.0-54
74.0-61
73.6-55
73.0-56
73.0-56
73.0-56
72.4-51
72.3-57
72.1-57
72.0-45
72.0-58
71.8-49
71.8-46
71.8-61
71.7-61
71.2-61
71.2-61
71.2-61
70.8-60
70.7-60
69.5-56
69.5-56
69.4-47
69.4-57
69.3-57
69.3-57
69.1-56
68.6-56
68.2-59
68.2+55
68.2-58
67.9-58
67.7-58
67.4-58
67.0-58
66.9-41
66.9-59
66.8-59
66.7-59
66.5-59
66.1-59
65.9-33
65.9-60
65.7-60
65.7-60
65.5-60
65.5-60
65.3-60
64.8-45
64.8-61
64.7-61
64.1-59
64.1+65
64.1-63
64.1-63
63.9-62
63.9-61
63.7-61
63.4-61
62.5-61
61.0-63
61.0-63
60.5-63
60.2-63
59.1-62
57.4-64
57.4-64
56.5-64
55.3-63
54.9-63
54.7-63
54.1-63
54.0-64
53.6-64
53.3-64
52.4-63
51.4-63
51.3-63
50.8-63
48.7-62
48.7-63
46.9-63
44.5-63
42.1-63
40.7-63
40.4-63
40.0-64
40.0-64
40.0-64
40.0-63
40.0-63
40.0-63
40.0-63
40.0-63
40.0-63
40.0-63
40.0-63
40.0-63
40.0-63
40.0-63
40.0-63
40.0-63
40.0-63
40.0-63
40.0-63
40.0-64
40.0-64
40.0-63
40.0-63
40.0-63
40.0-63
40.0-63
40.0-63
40.0-63
40.0-63
40.0-63
40.0-63
40.0-63
40.0-62
40.0-62
40.0-62
40.0-63
40.0-63
40.0-63
40.0-63
40.0-63
40.0-63
40.0-63
40.0-63
40.0-63
40.0-63
40.0-63
40.0-63
40.0-63
40.0-62
40.0-62
40.0-62
40.0-62
40.0-62
40.0-65
40.0-65
40.0-65
40.0-65
7

排行榜快照

本周综合评分前10的AI模型。

#模型评分7天变化
1Claude Fable 5Anthropic97.10
2Claude Fable 5 (batch)Anthropic97.1+377
3Claude Opus 5 (Fast)Anthropic95.1+171
4Claude Opus 5Anthropic95.1+171
5Claude Opus 4.8 (Fast)Anthropic95.1-1
6Claude Opus 4.8Anthropic95.1-1
7Claude Opus 4.7 (Fast)Anthropic95.1-5
8Claude Opus 4.7Anthropic95.1-5
9Claude Opus 4.7 (batch)Anthropic95.1+370
10Claude Opus 4.8 (batch)Anthropic94.6+369
8

本周数据

总排名数

300

平均评分

67.8

前10变动

9/10

最活跃的服务商

OpenAI

9

阅读更多分析

Frequently Asked Questions

The weekly report summarizes seven days of AI model ranking changes. It highlights the biggest gainers and losers by rank position, new models entering the leaderboard, models in a fragile watch-list state, and a snapshot of the current top 10. All data comes from hourly-updated composite scores across 290+ models.

Gainers and losers are ranked by their 7-day rank change, which measures how many positions a model moved up or down on the leaderboard over the past week. The models with the largest positive changes are the top gainers, and those with the largest negative changes are the top losers.

The watch list contains models currently in a "fragile" state, meaning their rankings are unstable and could degrade further. Monitoring these models is important if you rely on them in production, as they may experience significant quality or performance drops in the near term.

Weekly AI Performance Report - This Week in AI Models | LM Market Cap