Skip to content

LMC Capability Drift Tracker

How LLM capabilities spread through the catalog, quarter by quarter. Each curve shows the percentage of models released in a given release cohort that shipped with a given capability. 387 models tracked across 11 release quarters.

Catalog last refreshed 2026-08-28 (auto-updated hourly from live provider APIs).

Reasoning
92%
+34pp vs Y/Y
was 59% in 2025 Q3
Vision
60%
+26pp vs Y/Y
was 34% in 2025 Q3
Web Search
35%
+13pp vs Y/Y
was 22% in 2025 Q3
Function Calling
95%
+5pp vs Y/Y
was 90% in 2025 Q3
Image Output
0%
0pp vs Y/Y
was 0% in 2025 Q3
JSON Mode
81%
-4pp vs Y/Y
was 85% in 2025 Q3

Capability Adoption Curves

Percentage of models released in each quarter that ship with the given capability.

Reasoning
0% 92%
LMMarketCap.com
Vision
25% 60%
LMMarketCap.com
Function Calling
100% 95%
LMMarketCap.com
JSON Mode
75% 81%
LMMarketCap.com
Web Search
25% 35%
LMMarketCap.com
Image Output
0% 0%
LMMarketCap.com

Cohort Capability Matrix

Share of each release-quarter cohort that ships each capability, expressed as a percentage.

CohortNReasoningVisionFunction CallingJSON ModeWeb SearchImage Output
2024 Q140%25%100%75%25%0%
2024 Q290%33%44%56%0%22%
2024 Q3180%17%50%61%0%22%
2024 Q42015%20%55%35%5%10%
2025 Q12532%40%32%60%28%0%
2025 Q23168%68%90%77%48%0%
2025 Q34159%34%90%85%22%0%
2025 Q44770%70%83%89%40%9%
2026 Q15383%58%89%89%32%2%
2026 Q27097%79%93%93%46%6%
2026 Q36392%60%95%81%35%0%

Capability Debuts

First model in the catalog to ship each capability and how far the capability has spread since.

CapabilityFirst modelLatest adoption
Reasoning
Falcon3 10B Instruct92%
Vision
Claude 3 Haiku60%
Function Calling
GPT-3.5 Turbo95%
JSON Mode
GPT-3.5 Turbo81%
Web Search
Claude 3 Haiku35%
Image Output
DALL-E 30%

Latest models shipping reasoning

The fastest-spreading capability over the last year is reasoning. Here is what recently shipped with it.

ModelProviderReleased
Ling 3.0 Flash Fin (free)inclusionai2026-08-27
Qwen3.8 FlashAlibaba2026-08-26
GLM 5.3 FlashZhipu AI2026-08-26
Muse Spark 1.2 Contributormeta2026-08-21
DeepSeek V4 Flash Vision ExpDeepSeek2026-08-21
GLM Latest~z-ai2026-08-19
GLM 5.3Zhipu AI2026-08-18
Qwen3.8 27BAlibaba2026-08-14
Dots3-Note Preview (free)dots-studio2026-08-14
Gemini 3.7 FlashGoogle2026-08-13
Gemini 3.7 Flash (batch)Google2026-08-13
Seed 2.1 TurboByteDance2026-08-12

How the Capability Drift Tracker works

Every model we index has a release timestamp and a binary capability manifest. We bucket models into release-quarter cohorts using that timestamp, then compute the share of each cohort that shipped with a given capability. The result is an adoption curve per capability per quarter.

We exclude cohorts before 2024 Q1 from the visual curves because those quarters carry fewer than a dozen tracked releases each, so a single model can swing the percentage by tens of points. The cohort table below the charts shows all included quarters and their sample sizes.

Capability flags follow provider-declared manifests reconciled against documented API surfaces. A model that can accept images is counted under vision. A model that exposes a tool-calling endpoint is counted under function calling. A model that supports an explicit reasoning mode (visible or hidden thinking tokens) is counted under reasoning, even if the base mode is plain chat.

Frequently Asked Questions

It is a quarter-by-quarter view of how the LLM catalog has adopted each of six core capabilities: reasoning, vision, function calling, JSON mode, web search, and image output. Each cohort is defined by release quarter, and every model in that cohort contributes to the capability percentages for that quarter. This lets us see which capabilities are becoming table-stakes and which are still rare. Across the latest cohort (2026 Q3, N=63), 92% of new releases shipped with reasoning support and 60% shipped with vision input.

Provider roll-up hides the temporal story. A provider that shipped a reasoning model in 2024 Q4 and a non-reasoning model in 2025 Q2 would average out to 50% if we grouped by provider. Release-cohort grouping lets us see the calendar-quarter at which each capability crossed 50% adoption, which is a much more actionable question for buyers planning a rebuild.

Reasoning is the fastest-spreading capability over the last year. It was at 59% one year ago and now sits at 92% of new releases, a gain of 34 percentage points. Extended chain-of-thought / deliberate reasoning mode with visible or hidden thinking tokens.

Yes. The denominator for each cohort is every tracked model released in that quarter, including open source, free, and paid-only. We do not filter by license or pricing because capability adoption is a platform-wide question, and excluding open source would understate how quickly the open ecosystem is catching up on things like reasoning and vision.

Capability Drift Tracker - How LLM Capabilities Have Evolved | LM Market Cap