AI Model Pricing History
Track how AI model API pricing changes over time. Covering 405+ models across 63 providers, with weekly snapshots since 2026-02-16. For the dispersion view of how tightly frontier prices cluster, see the LMC Price Convergence Index.
Average Pricing by Provider (Current)
Average input and output cost per 1M tokens across each provider's paid models.
| Provider | Models | Avg Input $/M | Avg Output $/M |
|---|---|---|---|
| IBM | 2 | $0.034 | $0.106 |
| poolside | 2 | $0.075 | $0.150 |
| rekaai | 2 | $0.100 | $0.150 |
| ~deepseek | 1 | $0.080 | $0.160 |
| Meta | 8 | $0.138 | $0.313 |
| inclusionai | 4 | $0.045 | $0.336 |
| Microsoft | 2 | $0.345 | $0.380 |
| Tencent | 3 | $0.112 | $0.436 |
| Allen AI | 1 | $0.150 | $0.500 |
| Xiaomi | 2 | $0.287 | $0.575 |
| Upstage | 1 | $0.150 | $0.600 |
| StepFun | 2 | $0.150 | $0.725 |
| Inception | 1 | $0.250 | $0.750 |
| DeepSeek | 12 | $0.353 | $0.974 |
| ByteDance | 5 | $0.155 | $0.980 |
| arcee-ai | 2 | $0.485 | $1.02 |
| MiniMax | 9 | $0.286 | $1.18 |
| Meituan | 1 | $0.300 | $1.20 |
| Baidu | 1 | $0.420 | $1.25 |
| deepcogito | 1 | $1.25 | $1.25 |
Recent Price Drops (47 models)
Models that have decreased in price since tracking began.
| Model | Provider | Previous Output $/M | Current Output $/M | Change |
|---|---|---|---|---|
| GPT-5.6 Luna Pro | OpenAI | $6.00 | $0.600 | -90.0% |
| GPT-5.6 Luna | OpenAI | $6.00 | $0.600 | -90.0% |
| Ling-2.6-flash | inclusionai | $0.240 | $0.030 | -87.5% |
| MiMo-V2.5 | Xiaomi | $2.00 | $0.280 | -86.0% |
| Ling-2.6-1T | inclusionai | $2.50 | $0.625 | -75.0% |
| MiMo-V2.5-Pro | Xiaomi | $3.00 | $0.870 | -71.0% |
| DeepSeek V3.2 Speciale | DeepSeek | $1.20 | $0.431 | -64.1% |
| GPT-5.6 Terra Pro | OpenAI | $15.00 | $6.00 | -60.0% |
| GPT-5.6 Terra | OpenAI | $15.00 | $6.00 | -60.0% |
| Grok 4.20 | xAI | $6.00 | $2.50 | -58.3% |
| Grok 4.20 Multi-Agent | xAI | $6.00 | $2.50 | -58.3% |
| Llama Guard 3 8B | Meta | $0.060 | $0.030 | -50.0% |
| Kimi K2.6 | Moonshot AI | $4.66 | $2.44 | -47.6% |
| Qwen3.7 Max | Alibaba | $7.50 | $4.42 | -41.0% |
| GLM 5.2 | Zhipu AI | $4.00 | $2.42 | -39.5% |
| Qwen3 30B A3B Instruct 2507 | Alibaba | $0.300 | $0.193 | -35.6% |
| Qwen3 Max | Alibaba | $6.00 | $3.90 | -35.0% |
| Qwen-Plus | Alibaba | $1.20 | $0.780 | -35.0% |
| Qwen3.5-Flash | Alibaba | $0.400 | $0.260 | -35.0% |
| Qwen VL Max | Alibaba | $3.20 | $2.08 | -35.0% |
| Anthropic Claude Sonnet Latest | ~anthropic | $15.00 | $10.00 | -33.3% |
| MiniMax M2.5 | MiniMax | $1.20 | $0.900 | -25.0% |
| Mistral Nemo | Mistral AI | $0.040 | $0.030 | -25.0% |
| Qwen3.5 Plus 2026-04-20 | Alibaba | $2.40 | $1.80 | -25.0% |
| Qwen3.6 Flash | Alibaba | $1.50 | $1.13 | -25.0% |
Recent Price Increases (58 models)
| Model | Provider | Previous Output $/M | Current Output $/M | Change |
|---|---|---|---|---|
| Command R (08-2024) | Cohere | $0.600 | $10.00 | +1566.7% |
| Qwen3 30B A3B Thinking 2507 | Alibaba | $0.340 | $2.40 | +605.9% |
| Llama 3.2 11B Vision Instruct | Meta | $0.049 | $0.345 | +604.1% |
| Qwen3 235B A22B Instruct 2507 | Alibaba | $0.100 | $0.550 | +450.0% |
| Qwen2.5 Coder 32B Instruct | Alibaba | $0.200 | $1.00 | +400.0% |
| MoonshotAI Kimi Latest | ~moonshotai | $3.49 | $14.00 | +301.1% |
| Qwen3 235B A22B Thinking 2507 | Alibaba | $0.600 | $2.30 | +283.3% |
| Llama 3 8B Instruct | Meta | $0.040 | $0.140 | +250.0% |
| Gemma 3 27B | $0.150 | $0.450 | +200.0% | |
| Gemma 3n 4B | $0.040 | $0.120 | +200.0% | |
| Google Gemini Flash Latest | $3.00 | $7.50 | +150.0% | |
| Qwen3 VL 235B A22B Instruct | Alibaba | $0.880 | $1.90 | +115.9% |
| Qwen2.5 7B Instruct | Alibaba | $0.100 | $0.200 | +100.0% |
| Qwen3 30B A3B | Alibaba | $0.280 | $0.500 | +78.6% |
| Llama 3.1 8B Instruct | Meta | $0.050 | $0.080 | +60.0% |
| Qwen Plus 0728 (thinking) | Alibaba | $0.780 | $1.20 | +53.8% |
| Qwen3 VL 8B Thinking | Alibaba | $1.36 | $2.10 | +53.8% |
| DeepSeek V3 0324 | DeepSeek | $0.770 | $1.12 | +45.5% |
| QwQ 32B | Alibaba | $0.400 | $0.580 | +45.0% |
| Nemotron 3 Ultra | NVIDIA | $2.50 | $3.60 | +44.0% |
| Mistral Small 3.2 24B | Mistral AI | $0.180 | $0.250 | +38.9% |
| Kimi K2.5 | Moonshot AI | $2.20 | $2.85 | +29.5% |
| DeepSeek V3.1 | DeepSeek | $0.750 | $0.950 | +26.7% |
| DeepSeek V3.1 Terminus | DeepSeek | $0.790 | $1.00 | +26.6% |
| MiniMax M2.1 | MiniMax | $0.950 | $1.20 | +26.3% |
AI API Pricing Trends
AI model API pricing has been on a consistent downward trajectory since 2023. OpenAI's GPT-4 launched at $60/M output tokens; today, models with comparable capability cost under $5/M. This represents a 90%+ price reduction in under three years.
Key pricing trends observed across the industry:
- Race to the bottom: Competition between OpenAI, Anthropic, Google, and open-source providers drives continuous price cuts.
- Tiered pricing: Providers now offer multiple tiers (mini/flash models) at 10-50x lower cost than flagship models.
- Free tiers expanding: Google Gemini Flash, DeepSeek, and many open-source models available at zero cost through various providers.
- Caching discounts: Prompt caching can reduce input costs by 50-90% for repeated context.
Every Sunday at 23:00 UTC a cron job snapshots the full paid model catalog (currently 405 models across 63 providers) and writes a flat JSON file under data/weekly-snapshots/. The first snapshot in the table above was captured on 2026-02-16. Each row of every table on this page is sourced from those raw weekly files - no smoothing, no fills, no retroactive edits.
Most large providers adjust list pricing two to four times a year, typically alongside a new model generation. Mid-generation cuts of 50 to 90 percent are routine once a newer flagship ships at the same quality tier. The Recent Price Drops table above only counts changes greater than one percent, so cosmetic rounding does not show up as movement.
A handful of providers run experimental promotional pricing, particularly during model launches, and revert later. We treat every snapshot as authoritative for the week it was captured rather than smoothing those reversals away, so a temporary cut followed by a reversion shows as a drop in one week and an increase in another. This is intentional: it preserves the audit trail.
Google and DeepSeek have driven the steepest cuts in the dataset above, both by releasing flagship-tier reasoning models at fractions of incumbent pricing and by maintaining free tiers on Gemini Flash variants. Anthropic and OpenAI tend to cut older SKUs at the moment a new generation launches rather than adjusting the active flagship.
See /trackers/price-convergence-index. That page covers the LMC Price Convergence Index, which measures how tightly the frontier-tier price distribution clusters around its median in log space, with bias-corrected confidence intervals and a per-provider variance contribution breakdown. This page covers the per-model price trail: who moved when, and by how much.