Skip to content

Degradation Tracker

Detect when AI models may be declining. This tracker monitors rank movements (position changes on the leaderboard) and flags models that have dropped significantly over 24 hours or 7 days. A higher “degradation points” number means more warning signs.

Models at Risk

280

Declining (7d)

223

Fragile

201

Sustained Decline

34

Top Models by Degradation Risk Score

LMMarketCap.com

Models at Risk

280 models showing signs of degradation, ranked by risk score. Higher risk scores indicate more concerning performance trends.

Deg. PtsModelQualityRank 24hRank 7dSeverity
85Claude Opus 4.1Anthropic74.4-60-10high
37Claude Sonnet 4Anthropic73.9-4-14high
27GPT-4o (2024-08-06)OpenAI71.2+2-11high
27GPT-4oOpenAI71.2+2-11high
27GPT-4o (2024-05-13)OpenAI71.2+2-11high
27MiniMax M2MiniMax71.0+2-11high
27MiniMax M1MiniMax70.8+2-11high
27GLM 4.5 AirZhipu AI70.7+2-11high
27Llama 4 MaverickMeta70.7+2-11high
27Qwen3 30B A3BAlibaba64.1+2-11high
26Ling 3.0 Flash Sante (free)inclusionai40.0-1-10high
25Qwen3.6 Max PreviewAlibaba74.8+1-10high
25Hy3Tencent74.4+1-10high
25R1DeepSeek73.80-10high
25MiniMax M2.5MiniMax71.6+2-10high
24GPT-5 ProOpenAI88.7-1-9high
24GPT-5 Pro (batch)OpenAI88.7-1-9high
24GPT-5OpenAI88.7-1-9high
24GPT-5 (batch)OpenAI88.7-1-9high
24Gemini 3 Flash PreviewGoogle88.4-1-9high
24Gemini 3 Flash Preview (batch)Google88.4-1-9high
24Muse Spark 1.3 Contributormeta40.0-1-9high
23o1OpenAI74.4+1-9high
23o3 Mini (batch)OpenAI74.3+1-9high
23Qwen3.6 PlusAlibaba74.1+1-9high
23MiniMax M3MiniMax73.9+1-9high
23o1-proOpenAI73.6+1-9high
23Qwen3.8 27BAlibaba73.4+1-9high
23MiMo-V2.5Xiaomi73.0+1-9high
23Gemma 4 26B A4B Google73.0+1-9high
23Gemma 4 26B A4B (free)Google73.0+1-9high
23Qwen3.8 27B (free)Alibaba72.9+1-9high
23Inklingthinkingmachines72.9+1-9high
23Inkling (free)thinkingmachines72.9+1-9high
23DeepSeek V3 0324DeepSeek71.8+2-9high
23Mistral Medium 3.5Mistral AI71.6+2-9high
23Mistral LargeMistral AI65.9+2-9high
23Composer 2Cursor65.7+2-9high
23Composer 2 FastCursor65.7+2-9high
23GLM 4.6VZhipu AI65.6+2-9high
23Qwen3 235B A22B Thinking 2507Alibaba65.3+2-9high
23Llama 3.1 70B InstructMeta65.3+2-9high
23GPT-4OpenAI64.9+2-9high
23Qwen3 235B A22B Instruct 2507Alibaba64.7+2-9high
23o3 Mini HighOpenAI63.9+2-9high
23GLM 4.7 FlashZhipu AI63.5+2-9high
23Mixtral 8x22B InstructMistral AI63.4+2-9high
23Trinity Large Thinkingarcee-ai63.1+2-9high
23GPT-4o-mini (batch)OpenAI62.5+2-9high
23GLM 4.5VZhipu AI62.3+2-9high
23GPT-5 NanoOpenAI61.9+2-9high
23gpt-oss-120bOpenAI61.6+2-9high
23GPT-4o-mini (2024-07-18)OpenAI56.5+2-9high
23GPT-4.1 MiniOpenAI56.2+2-9high
23Ternary Bonsai 2 27Bprism-ml40.00-9high
23Paretounbiased40.00-9high
23DeepSeek Pro Latest~deepseek40.00-9high
23DeepSeek Flash Latest~deepseek40.00-9high
23Schematron V2 Turboinference-net40.00-9high
23Schematron V2 Smallinference-net40.00-9high
23GPT Astra Latest~openai40.00-9high
23GPT Sol Latest~openai40.00-9high
23GPT Terra Latest~openai40.00-9high
23GPT Luna Latest~openai40.00-9high
23Fugu Ultra v2sakana40.00-9high
23Fugu Maxsakana40.00-9high
23Ling 3.0 Flash VLinclusionai40.00-9high
23Ling 3.0 Flash VL (free)inclusionai40.00-9high
23DeepSeek V4.1 FlashDeepSeek40.00-9high
22Grok 4.20 Multi-AgentxAI87.9-1-8high
22Hy4 previewTencent40.0-1-8high
22Ling 3.0 Flash Fininclusionai40.0-1-8high
22Ling 3.0 Flash Fin (free)inclusionai40.0-1-8high
22GLM Flash Latest~z-ai40.0-1-8high
22Qwen3.8 FlashAlibaba40.0-1-8high
22Muse Spark 1.2 Contributormeta40.0-1-8high
22DeepSeek V4 Flash Vision ExpDeepSeek40.0-1-8high
21GLM 5Zhipu AI78.0+2-8high
21GLM 5.3 FlashXZhipu AI77.9+2-8high
21GLM 5.3 Flash (batch)Zhipu AI77.9+2-8high
21GLM 4.5Zhipu AI75.1+1-8high
21Claude Haiku 4.5Anthropic69.9+2-8high
21Claude Haiku 4.5 (batch)Anthropic69.9+2-8high
21Qwen3 Max ThinkingAlibaba68.2+2-8high
21MiniMax M2-herMiniMax68.1+2-8high
21GPT-4.1OpenAI67.7+2-8high
21GPT-4.1 (batch)OpenAI67.7+2-8high
21Qwen3 Next 80B A3B InstructAlibaba66.7+2-8high
21GPT-4 TurboOpenAI66.7+2-8high
21Qwen3.5-9BAlibaba66.5+2-8high
21Step 3.5 FlashStepFun66.1+2-8high
21Qwen3 30B A3B Thinking 2507Alibaba64.1+2-8high
21GPT-4 Turbo (batch)OpenAI61.5+3-8high
21Mercury 2.5Inception61.2+3-8high
21Qwen3 8BAlibaba61.0+3-8high
21Mercury 2Inception60.9+3-8high
21Nova 2 LiteAmazon60.4+3-8high
21Llama 4 ScoutMeta60.2+3-8high
21Phi 4Microsoft60.2+3-8high
21GPT-4.1 Mini (batch)OpenAI58.9+3-8high
21GPT-4.1 Nano (batch)OpenAI58.9+3-8high
21gpt-oss-20bOpenAI57.4+3-8high
21Qwen3 235B A22BAlibaba54.0+2-8high
21Granite 4.2 8BIBM53.8+2-8high
21GPT-5 MiniOpenAI53.8+2-8high
20Grok 4.20xAI88.8-1-7high
19GLM 5.3 (batch)Zhipu AI78.6+1-7high
19GLM 5.2Zhipu AI78.6+1-7high
19Qwen3.5-122B-A10BAlibaba77.7+2-7high
19Gemma 2 27BGoogle77.4+2-7high
19Qwen3.7 PlusAlibaba75.6+1-7high
19GLM 5V TurboZhipu AI72.3+2-7high
19o4 Mini HighOpenAI72.1+2-7high
19GPT-4o (batch)OpenAI72.0+2-7high
19DeepSeek V3DeepSeek69.5+2-7high
19Qwen3 VL 235B A22B InstructAlibaba69.3+2-7high
19DeepSeek V3.1 TerminusDeepSeek69.3+2-7high
19GPT-4o-miniOpenAI69.3+2-7high
19Inkling Smallthinkingmachines68.7+2-7high
19Inkling Small (free)thinkingmachines68.7+2-7high
19Qwen3.5-FlashAlibaba68.6+2-7high
19Hy3 previewTencent68.4+2-7high
19Qwen3 MaxAlibaba67.4+2-7high
19Mistral Large 3 2512 (batch)Mistral AI67.0+2-7high
19Llama 3.3 70B InstructMeta66.8+2-7high
19Claude 3 HaikuAnthropic51.3+2-7high
19Command ACohere50.8+2-7high
19Command R+ (08-2024)Cohere48.7+2-7high
19Kimi K2.5Moonshot AI47.5+2-7high
19Llama 3.1 8B InstructMeta44.5+2-7high
19GPT-4.1 NanoOpenAI42.1+2-7high
19R1 Distill Llama 70BDeepSeek40.7+2-7high
19Hy-MT2-1.8BTencent40.00-7high
19Hy-MT2-30B-A3BTencent40.00-7high
19GLM Latest~z-ai40.00-7high
19Hy-MT2-7BTencent40.00-7high
19Dots3-Note Preview (free)dots-studio40.00-7high
19Gemini 3.7 FlashGoogle40.00-7high
19Gemini 3.7 Flash (batch)Google40.00-7high
19Seed 2.1 TurboByteDance40.00-7high
19Qwen3.8 2.4T A95BAlibaba40.00-7high
18GPT-5.1 (batch)OpenAI87.8-1-6high
18Claude Sonnet 4.6Anthropic85.2-1-6high
18Claude Sonnet 4.6 (batch)Anthropic85.2-1-6high
18Claude Opus 4.5Anthropic85.1-1-6high
18Claude Opus 4.5 (batch)Anthropic85.1-1-6high
18Gemini 2.5 ProGoogle83.5-1-6high
18Gemini 2.5 Pro (batch)Google83.5-1-6high
18Gemini 2.5 Pro Preview 06-05Google83.5-1-6high
18DeepSeek V3.2DeepSeek83.4-1-6high
18Claude Sonnet 4.5Anthropic82.4-1-6high
18Claude Sonnet 4.5 (batch)Anthropic82.4-1-6high
17Muse Spark 1.1meta80.9+1-6high
17Gemma 4 31BGoogle80.5+1-6high
17Gemma 4 31B (free)Google80.5+1-6high
17Qwen3.5 397B A17BAlibaba79.4+1-6high
17R1 0528DeepSeek79.4+1-6high
17GPT-5.4 NanoOpenAI79.3+1-6high
17GPT-5.4 Nano (batch)OpenAI79.3+1-6high
17GPT-5.4 MiniOpenAI79.3+1-6high
17GPT-5.4 Mini (batch)OpenAI79.3+1-6high
17Gemini 3.1 Flash Lite PreviewGoogle79.3+1-6high
17Gemini 2.5 Flash LiteGoogle79.1+1-6high
17Gemini 2.5 Flash Lite (batch)Google79.1+1-6high
17Gemini 2.5 FlashGoogle79.1+1-6high
17Gemini 2.5 Flash (batch)Google79.1+1-6high
17Gemini 3.5 FlashGoogle79.0+1-6high
17Gemini 3.5 Flash (batch)Google79.0+1-6high
17GLM 5.3 FlashZhipu AI78.0+2-6high
17MiMo-V2.5-ProXiaomi76.2+1-6high
17Qwen3.5-35B-A3BAlibaba76.0+1-6high
17GLM 5.2 (free)Zhipu AI75.7+1-6high
17Kimi K2.6Moonshot AI75.7+1-6high
17o3 MiniOpenAI75.3+1-6high
17Claude Opus 4.1 (batch)Anthropic75.2+1-6high
17Command R (08-2024)Cohere48.7+3-6high
17Seed-2.0-CodeByteDance40.0+1-6high
17DeepSeek V4 Pro 0813DeepSeek40.0+1-6high
11GPT-5.1-Codex-MiniOpenAI87.8-1-5high
11o3 ProOpenAI86.7-1-5high
11o3OpenAI86.7-1-5high
11o3 (batch)OpenAI86.7-1-5high
10GPT-6 AstraOpenAI81.8+1-5high
10GPT-6 Astra (batch)OpenAI81.8+1-5high
10GPT-6 Astra ProOpenAI81.8+1-5high
10GPT-6 Astra Pro (batch)OpenAI81.8+1-5high
10o4 MiniOpenAI81.4+1-5high
10o4 Mini (batch)OpenAI81.4+1-5high
10Gemini 3.8 FlashGoogle81.0+1-5high
10Gemini 3.8 Flash (batch)Google81.0+1-5high
10Qwen3.5-27BAlibaba77.0+2-5high
10GPT-5 Mini (batch)OpenAI76.8+2-5high
10GPT-5 Nano (batch)OpenAI76.8+2-5high
10Gemini 3.5 Flash LiteGoogle76.5+2-5high
10Gemini 3.5 Flash Lite (batch)Google76.5+2-5high
10LFM2.5-2.6B (free)Liquid AI40.0+2-5high
10Nemotron 3.5 LightningNVIDIA40.0+2-5high
10Nemotron 3.5 Lightning (free)NVIDIA40.0+2-5high
10Sakana Namazusakana40.0+2-5high
10Solar Pro 4Upstage40.0+2-5high
10Muse Glimmer 30Bmeta40.0+2-5high
8Grok 4.3 (batch)xAI81.0+1-4medium
8MiniMax M2.1MiniMax71.0+2-4medium
6GPT-5.2 ProOpenAI90.50-3medium
6GPT-5.2 Pro (batch)OpenAI90.50-3medium
6GPT-5.2OpenAI90.50-3medium
6GPT-5.2 (batch)OpenAI90.50-3medium
6Claude Opus 4.6Anthropic90.40-3medium
6Claude Opus 4.6 (batch)Anthropic90.40-3medium
6GPT-5.6 Luna ProOpenAI89.00-3medium
6GPT-5.6 Luna Pro (batch)OpenAI89.00-3medium
6GPT-5.6 LunaOpenAI89.00-3medium
6GPT-5.6 Luna (batch)OpenAI89.00-3medium
6GPT-5.6 Terra ProOpenAI89.00-3medium
6GPT-5.6 Terra Pro (batch)OpenAI89.00-3medium
6GPT-5.6 TerraOpenAI89.00-3medium
6GPT-5.6 Terra (batch)OpenAI89.00-3medium
6GPT-5.6 Sol ProOpenAI89.00-3medium
6GPT-5.6 Sol Pro (batch)OpenAI89.00-3medium
6GPT-5.6 SolOpenAI89.00-3medium
6GPT-5.6 Sol (batch)OpenAI89.00-3medium
6Grok 4.5xAI88.8-1+57medium
6Grok 4.3xAI88.8-1+31medium
6Claude Sonnet 5Anthropic85.2-1+232medium
6DeepSeek V4 Flash Latest~deepseek40.0+3-3medium
6DeepSeek V4 Flash 0731DeepSeek40.0+3-3medium
5Claude Opus 5Anthropic95.10+273medium
5GPT-5.5 ProOpenAI92.70+9medium
5GPT-5.3-CodexOpenAI90.50+19medium
5GPT-5.2 ChatOpenAI90.50+81medium
5Muse Spark 1.3meta80.9+1+162medium
5Muse Spark 1.2meta80.9+1+190medium
5GLM 5.1Zhipu AI78.0+2+12medium
5GLM 5 TurboZhipu AI78.0+2+44medium
5Qwen3.8 Max (0902)Alibaba75.6+1+121medium
5GLM 4.7Zhipu AI75.1+1+14medium
5GLM 4.6Zhipu AI75.1+1+28medium
5Qwen3.7 MaxAlibaba74.8+1+177medium
5Qwen3.5 Plus 2026-04-20Alibaba74.8+1+188medium
5DeepSeek V3.1DeepSeek71.8+2+8medium
5MiniMax M2.7MiniMax71.6+2+7medium
5GPT-4o (2024-11-20)OpenAI71.2+2+61medium
5Qwen3.5 Plus 2026-02-15Alibaba68.2+2+160medium
5Qwen3 Next 80B A3B ThinkingAlibaba66.7+2+7medium
5Mistral Large 2407Mistral AI65.9+2+24medium
5Qwen3 30B A3B Instruct 2507Alibaba64.1+2+171medium
4Claude Opus 4.7Anthropic95.10-2low
4Claude Opus 4.7 (batch)Anthropic95.10-2low
4GPT-5.5 Pro (batch)OpenAI92.70-2low
4GPT-5.5OpenAI92.70-2low
4GPT-5.5 (batch)OpenAI92.70-2low
4Gemini 3.1 Pro Preview Custom ToolsGoogle92.20-2low
4Gemini 3.1 Pro PreviewGoogle92.20-2low
4Gemini 3.1 Pro Preview (batch)Google92.20-2low
4GPT-5.4 ProOpenAI91.90-2low
4GPT-5.4 Pro (batch)OpenAI91.90-2low
4GPT-5.4OpenAI91.90-2low
4GPT-5.4 (batch)OpenAI91.90-2low
4GPT-5.2-CodexOpenAI90.50-2low
4Qwen3.7 FlashAlibaba40.0+4-2low
2Claude Opus 4.8 (batch)Anthropic94.60-1low
2Claude Opus 5 (batch)Anthropic40.0+4-1low
2Ling 3.0 Flashinclusionai40.0+4-1low
2Laguna S 2.1poolside40.0+4-1low
2Laguna S 2.1 (free)poolside40.0+4-1low
2Gemini 3.6 FlashGoogle40.0+4-1low
2Gemini 3.6 Flash (batch)Google40.0+4-1low
2LongCat 2.0Meituan40.0+4-1low
2Kimi K3Moonshot AI40.0+4-1low
2Kimi K3 (batch)Moonshot AI40.0+4-1low
2KAT-Coder-Pro V2.5Kuaishou40.0+4-1low
2Grok Latest~x-ai40.0+4-1low
2Aion-3.0-Miniaion-labs40.0+4-1low
2Aion-3.0aion-labs40.0+4-1low
2Laguna XS 2.1poolside40.0+4-1low
2Laguna XS 2.1 (free)poolside40.0+4-1low
1Grok 4.6xAI88.8-1+4low
1GPT-5.1-Codex-MaxOpenAI88.7-1+2low
1GPT-5.1OpenAI88.7-1+2low
1GPT-5.1-CodexOpenAI88.7-1+3low

Stable Models

14 models with no decline and a stable ranking state. These models are performing consistently.

#ModelScore24h7dState
1Claude Fable 5.1Anthropic95.900stable
2Claude Fable 5.1 (batch)Anthropic95.900stable
3Claude Fable 5Anthropic95.900stable
4Claude Fable 5 (batch)Anthropic95.900stable
6Claude Opus 4.8Anthropic95.10+1stable
100GLM 5.3Zhipu AI78.6+1+5stable
150DeepSeek V3.2 ExpDeepSeek71.8+2+5stable
168Qwen3 VL 235B A22B ThinkingAlibaba69.3+2+5stable
295Claude Sonnet 5 (batch)Anthropic40.0+40stable
296Fugu Ultrasakana40.0+40stable
297North Mini Code (free)Cohere40.0+40stable
298Kimi K2.7 CodeMoonshot AI40.0+40stable
299Claude Fable Latest~anthropic40.0+40stable
300Nemotron 3.5 Content SafetyNVIDIA40.0+40stable

How Degradation Is Detected

Our degradation detection system uses multiple signals to identify models that may be declining in quality or reliability.

Declining (7d)

Models whose 7-day rank change is worse than -2 positions. A sustained drop of more than two ranks over a week suggests the model may be losing ground to competitors or experiencing performance issues.

Fragile State

Models classified as "fragile" by our scoring system. These models have inconsistent performance metrics or borderline scores that could shift significantly with small changes in evaluation data.

Sustained Decline

Models declining on both the 24-hour and 7-day timeframes. When a model is losing rank on both short and medium-term windows, it indicates a persistent downward trend rather than temporary fluctuation.

Risk Score

The degradation risk score combines multiple signals: 7-day rank decline weighted 2x, 24-hour rank decline weighted 1x, plus 5 bonus points for fragile state. Higher scores indicate greater risk of meaningful performance degradation.

Related

Frequently Asked Questions

The tracker uses a multi-signal approach: it monitors 7-day rank decline (weighted 2x), 24-hour rank drops (weighted 1x), and fragile state classification (+5 points). Models are scored on a degradation risk scale where higher values indicate more warning signs of performance decline.

A fragile state indicates that a model has inconsistent performance metrics or borderline scores that could shift significantly with small changes in evaluation data. Fragile models are at higher risk of further ranking drops and warrant closer monitoring.

Yes, models can recover. Degradation may be temporary due to API issues, benchmark fluctuations, or scoring recalibrations. Models that show sustained decline over multiple weeks are more concerning than those with short-term dips. The tracker monitors both 24-hour and 7-day windows to help distinguish temporary noise from real trends.

AI Model Degradation Tracker - Detect Performance Drops | LM Market Cap