Skip to content

Large Context Window AI Models

Compare 429 AI models with context windows of 32K tokens or more. The largest models support 2M tokens -- enough to process entire codebases, books, or hundreds of documents in a single prompt. Data updated hourly.

170
1M+ Tokens
153
200K+ Tokens
79
128K Tokens
27
32K–64K Tokens
2M
Largest Context
429
Models (32K+)
548K
Average Context

1M+ Tokens

(170 models)

The largest context windows available. These models can process entire codebases, full books, or hundreds of documents in a single request.

#ModelProviderContext WindowPagesInput / 1MOutput / 1M
1Grok 4.20xAI2M~3K$1.25$2.50
2Grok 4.20 Multi-AgentxAI2M~3K$1.25$2.50
3DeepSeek V4 Flash 0731DeepSeek1.3M~2K$0.03$0.32
4DeepSeek V4 Flash Latest~deepseek1.3M~2K$0.03$0.32
5GLM 5.3Zhipu AI1.3M~2K$1.40$4.40
6GLM 5.3 FlashZhipu AI1.3M~2K$0.04$0.60
7GLM Flash Latest~z-ai1.3M~2K$0.04$0.14
8GLM Latest~z-ai1.3M~2K$0.56$1.76
9Llama 4 ScoutMeta1.3M~2K$0.10$0.30
10GPT Astra Latest~openai1.1M~1K$10.00$50.00
11GPT Luna Latest~openai1.1M~1K$0.10$0.50
12GPT Sol Latest~openai1.1M~1K$2.00$10.00
13GPT Terra Latest~openai1.1M~1K$2.00$12.00
14GPT-5.4OpenAI1.1M~1K$2.50$15.00
15GPT-5.4 (batch)OpenAI1.1M~1K$1.25$7.50
16GPT-5.4 ProOpenAI1.1M~1K$30.00$180.00
17GPT-5.4 Pro (batch)OpenAI1.1M~1K$15.00$90.00
18GPT-5.5OpenAI1.1M~1K$5.00$30.00
19GPT-5.5 (batch)OpenAI1.1M~1K$2.50$15.00
20GPT-5.5 ProOpenAI1.1M~1K$30.00$180.00
21GPT-5.5 Pro (batch)OpenAI1.1M~1K$15.00$90.00
22GPT-5.6 LunaOpenAI1.1M~1K$0.20$1.20
23GPT-5.6 Luna (batch)OpenAI1.1M~1K$0.10$0.60
24GPT-5.6 Luna ProOpenAI1.1M~1K$0.20$1.20
25GPT-5.6 Luna Pro (batch)OpenAI1.1M~1K$0.10$0.60
26GPT-5.6 SolOpenAI1.1M~1K$2.00$10.00
27GPT-5.6 Sol (batch)OpenAI1.1M~1K$1.00$5.00
28GPT-5.6 Sol ProOpenAI1.1M~1K$2.00$10.00
29GPT-5.6 Sol Pro (batch)OpenAI1.1M~1K$1.00$5.00
30GPT-5.6 TerraOpenAI1.1M~1K$2.00$12.00
31GPT-5.6 Terra (batch)OpenAI1.1M~1K$1.00$6.00
32GPT-5.6 Terra ProOpenAI1.1M~1K$2.00$12.00
33GPT-5.6 Terra Pro (batch)OpenAI1.1M~1K$1.00$6.00
34GPT-6 AstraOpenAI1.1M~1K$10.00$50.00
35GPT-6 Astra (batch)OpenAI1.1M~1K$5.00$25.00
36GPT-6 Astra ProOpenAI1.1M~1K$10.00$50.00
37GPT-6 Astra Pro (batch)OpenAI1.1M~1K$5.00$25.00
38GPT-6 LunaOpenAI1.1M~1K$0.10$0.50
39GPT-6 Luna (batch)OpenAI1.1M~1K$0.05$0.25
40GPT-6 Luna ProOpenAI1.1M~1K$0.10$0.50
41GPT-6 Luna Pro (batch)OpenAI1.1M~1K$0.05$0.25
42GPT-6 SolOpenAI1.1M~1K$2.00$10.00
43GPT-6 Sol (batch)OpenAI1.1M~1K$1.00$5.00
44GPT-6 Sol ProOpenAI1.1M~1K$2.00$10.00
45GPT-6 Sol Pro (batch)OpenAI1.1M~1K$1.00$5.00
46MiMo-V2.5Xiaomi1.1M~1K$0.14$0.28
47MiMo-V2.5-ProXiaomi1.1M~1K$0.43$0.87
48LongCat 2.0Meituan1.0M~1K$0.30$1.20
49DeepSeek Flash Latest~deepseek1.0M~1K$0.04$0.49
50DeepSeek Pro Latest~deepseek1.0M~1K$0.35$1.05
51DeepSeek V4 Flash 0423DeepSeek1.0M~1K$0.05$0.10
52DeepSeek V4 Flash Vision ExpDeepSeek1.0M~1K$0.22$0.66
53DeepSeek V4 Pro 0423DeepSeek1.0M~1K$0.82$1.64
54DeepSeek V4 Pro 0813DeepSeek1.0M~1K$0.46$1.39
55DeepSeek V4.1 FlashDeepSeek1.0M~1K$0.30$1.20
56DeepSeek V4.1 Flash (batch)DeepSeek1.0M~1K$0.11$0.34
57Ember-1fireworks1.0M~1K$3.00$15.00
58Gemini 2.5 FlashGoogle1.0M~1K$0.30$2.50
59Gemini 2.5 Flash (batch)Google1.0M~1K$0.15$1.25
60Gemini 2.5 Flash LiteGoogle1.0M~1K$0.10$0.40
61Gemini 2.5 Flash Lite (batch)Google1.0M~1K$0.05$0.20
62Gemini 2.5 ProGoogle1.0M~1K$1.25$10.00
63Gemini 2.5 Pro (batch)Google1.0M~1K$0.63$5.00
64Gemini 2.5 Pro Preview 06-05Google1.0M~1K$1.25$10.00
65Gemini 3 Flash PreviewGoogle1.0M~1K$0.50$3.00
66Gemini 3 Flash Preview (batch)Google1.0M~1K$0.25$1.50
67Gemini 3.1 Flash LiteGoogle1.0M~1K$0.25$1.50
68Gemini 3.1 Flash Lite (batch)Google1.0M~1K$0.13$0.75
69Gemini 3.1 Flash Lite PreviewGoogle1.0M~1K$0.25$1.50
70Gemini 3.1 Pro PreviewGoogle1.0M~1K$2.00$12.00
71Gemini 3.1 Pro Preview (batch)Google1.0M~1K$1.00$6.00
72Gemini 3.1 Pro Preview Custom ToolsGoogle1.0M~1K$2.00$12.00
73Gemini 3.5 FlashGoogle1.0M~1K$1.50$9.00
74Gemini 3.5 Flash (batch)Google1.0M~1K$0.75$4.50
75Gemini 3.5 Flash LiteGoogle1.0M~1K$0.30$2.50
76Gemini 3.5 Flash Lite (batch)Google1.0M~1K$0.15$1.25
77Gemini 3.6 FlashGoogle1.0M~1K$0.75$3.75
78Gemini 3.6 Flash (batch)Google1.0M~1K$0.38$1.88
79Gemini 3.7 FlashGoogle1.0M~1K$0.75$3.75
80Gemini 3.7 Flash (batch)Google1.0M~1K$0.38$1.88
81Gemini 3.8 FlashGoogle1.0M~1K$0.75$3.75
82Gemini 3.8 Flash (batch)Google1.0M~1K$0.38$1.88
83Gemini Flash Latest~google1.0M~1K$0.75$3.75
84Gemini Pro Latest~google1.0M~1K$2.00$12.00
85GLM 5.2Zhipu AI1.0M~1K$0.65$2.04
86GLM 5.3 (batch)Zhipu AI1.0M~1K$0.45$2.00
87GLM 5.3 Flash (batch)Zhipu AI1.0M~1K$0.06$0.20
88GLM 5.3 FlashXZhipu AI1.0M~1K$0.37$1.25
89Hy4 previewTencent1.0M~1K$0.83$2.50
90Inklingthinkingmachines1.0M~1K$1.00$4.05
91Inkling (free)thinkingmachines1.0M~1KFreeFree
92Inkling Smallthinkingmachines1.0M~1K$0.45$1.20
93Inkling Small (free)thinkingmachines1.0M~1KFreeFree
94Kimi K3Moonshot AI1.0M~1K$0.88$10.53
95Kimi K3 (batch)Moonshot AI1.0M~1K$2.28$11.40
96Kimi Latest~moonshotai1.0M~1K$0.88$10.53
97Laguna S 2.1poolside1.0M~1K$0.09$0.18
98Llama 4 MaverickMeta1.0M~1K$0.19$0.65
99Lyria 3 Clip PreviewGoogle1.0M~1KFreeFree
100Lyria 3 Pro PreviewGoogle1.0M~1KFreeFree
101MiMo-V2.6-FlashXiaomi1.0M~1K$0.14$0.28
102MiMo-V2.6-ProXiaomi1.0M~1K$0.43$0.87
103MiMo-V2.6-Pro-UltraSpeedXiaomi1.0M~1K$4.35$8.70
104MiniMax M3MiniMax1.0M~1K$0.30$1.20
105Muse Spark 1.1meta1.0M~1K$1.25$4.25
106Muse Spark 1.2meta1.0M~1K$1.25$4.25
107Muse Spark 1.2 Contributormeta1.0M~1K$0.10$0.20
108Muse Spark 1.3meta1.0M~1K$1.25$4.25
109Muse Spark 1.3 Contributormeta1.0M~1K$0.10$0.20
110Qwen3.8 2.4T A95BAlibaba1.0M~1K$2.00$6.00
111GPT-4.1OpenAI1.0M~1K$2.00$8.00
112GPT-4.1 (batch)OpenAI1.0M~1K$1.00$4.00
113GPT-4.1 MiniOpenAI1.0M~1K$0.40$1.60
114GPT-4.1 Mini (batch)OpenAI1.0M~1K$0.20$0.80
115GPT-4.1 NanoOpenAI1.0M~1K$0.10$0.40
116GPT-4.1 Nano (batch)OpenAI1.0M~1K$0.05$0.20
117Palmyra X5Writer1.0M~1K$0.60$6.00
118MiniMax-01MiniMax1.0M~1K$0.20$1.10
119Claude Fable 5Anthropic1M~1K$10.00$50.00
120Claude Fable 5 (batch)Anthropic1M~1K$5.00$25.00
121Claude Fable 5.1Anthropic1M~1K$10.00$50.00
122Claude Fable 5.1 (batch)Anthropic1M~1K$5.00$25.00
123Claude Fable Latest~anthropic1M~1K$10.00$50.00
124Claude Opus 4.6Anthropic1M~1K$5.00$25.00
125Claude Opus 4.6 (batch)Anthropic1M~1K$2.50$12.50
126Claude Opus 4.7Anthropic1M~1K$5.00$25.00
127Claude Opus 4.7 (batch)Anthropic1M~1K$2.50$12.50
128Claude Opus 4.8Anthropic1M~1K$5.00$25.00
129Claude Opus 4.8 (batch)Anthropic1M~1K$2.50$12.50
130Claude Opus 5Anthropic1M~1K$5.00$25.00
131Claude Opus 5 (batch)Anthropic1M~1K$2.50$12.50
132Claude Opus 5.5Anthropic1M~1K$4.00$20.00
133Claude Opus 5.5 (batch)Anthropic1M~1K$2.00$10.00
134Claude Opus Latest~anthropic1M~1K$4.00$20.00
135Claude Sonnet 4.5Anthropic1M~1K$3.00$15.00
136Claude Sonnet 4.5 (batch)Anthropic1M~1K$1.50$7.50
137Claude Sonnet 4.6Anthropic1M~1K$3.00$15.00
138Claude Sonnet 4.6 (batch)Anthropic1M~1K$1.50$7.50
139Claude Sonnet 5Anthropic1M~1K$2.00$10.00
140Claude Sonnet 5 (batch)Anthropic1M~1K$1.00$5.00
141Claude Sonnet Latest~anthropic1M~1K$2.00$10.00
142Fugu Maxsakana1M~1K$2.00$6.00
143Fugu Ultrasakana1M~1K$5.00$30.00
144Fugu Ultra v2sakana1M~1K$5.00$30.00
145GLM 5.3 PrimeZhipu AI1M~1K$2.80$8.80
146Grok 4.3xAI1M~1K$1.25$2.50
147Grok 4.3 (batch)xAI1M~1K$1.00$2.00
148MiniMax M1MiniMax1M~1K$0.40$2.20
149Nemotron 3 Ultra (free)NVIDIA1M~1KFreeFree
150Nemotron 3.5 Lightning (free)NVIDIA1M~1KFreeFree
151Nova 2 LiteAmazon1M~1K$0.30$2.50
152Nova Premier 1.0Amazon1M~1K$2.50$12.50
153Qwen Plus 0728Alibaba1M~1K$0.26$0.78
154Qwen-PlusAlibaba1M~1K$0.26$0.78
155Qwen3 Coder FlashAlibaba1M~1K$0.20$0.97
156Qwen3 Coder PlusAlibaba1M~1K$0.65$3.25
157Qwen3.5 Plus 2026-02-15Alibaba1M~1K$0.26$1.56
158Qwen3.5 Plus 2026-04-20Alibaba1M~1K$0.30$1.80
159Qwen3.5-FlashAlibaba1M~1K$0.07$0.26
160Qwen3.6 FlashAlibaba1M~1K$0.19$1.13
161Qwen3.6 PlusAlibaba1M~1K$0.33$1.95
162Qwen3.7 FlashAlibaba1M~1K$0.03$0.13
163Qwen3.7 MaxAlibaba1M~1K$1.48$4.42
164Qwen3.7 PlusAlibaba1M~1K$0.32$1.28
165Qwen3.8 27BAlibaba1M~1K$0.42$3.00
166Qwen3.8 FlashAlibaba1M~1K$0.15$0.47
167Qwen3.8 Max (0902)Alibaba1M~1K$2.00$6.00
168Qwen3.8 Max PrimeAlibaba1M~1K$4.00$12.00
169Qwen3.8 Omni FlashAlibaba1M~1K$0.15$0.47
170Space Bunny Alphastealth1M~1KFreeFree

200K+ Tokens

(153 models)

Extended context models ideal for long documents, legal contracts, research papers, and multi-file code analysis.

#ModelProviderContext WindowPagesInput / 1MOutput / 1M
1Solar Mini 4Upstage524K~699$0.05$0.20
2Solar Pro 4Upstage524K~699$0.09$0.36
3Dots3-Note Preview (free)dots-studio512K~683FreeFree
4Grok 4.5xAI500K~667$2.00$6.00
5Grok 4.6xAI500K~667$2.00$6.00
6Grok 4.7xAI500K~667$1.60$4.80
7Grok Latest~x-ai500K~667$1.60$4.80
8GPT Chat LatestOpenAI400K~533$5.00$30.00
9GPT Mini Latest~openai400K~533$0.75$4.50
10GPT-5OpenAI400K~533$1.25$10.00
11GPT-5 (batch)OpenAI400K~533$0.63$5.00
12GPT-5 ImageOpenAI400K~533$10.00$10.00
13GPT-5 Image MiniOpenAI400K~533$2.50$2.00
14GPT-5 MiniOpenAI400K~533$0.25$2.00
15GPT-5 Mini (batch)OpenAI400K~533$0.13$1.00
16GPT-5 NanoOpenAI400K~533$0.05$0.40
17GPT-5 Nano (batch)OpenAI400K~533$0.02$0.20
18GPT-5 ProOpenAI400K~533$15.00$120.00
19GPT-5 Pro (batch)OpenAI400K~533$7.50$60.00
20GPT-5.1OpenAI400K~533$1.25$10.00
21GPT-5.1 (batch)OpenAI400K~533$0.63$5.00
22GPT-5.1-CodexOpenAI400K~533$1.25$10.00
23GPT-5.1-Codex-MaxOpenAI400K~533$1.25$10.00
24GPT-5.1-Codex-MiniOpenAI400K~533$0.25$2.00
25GPT-5.2OpenAI400K~533$1.75$14.00
26GPT-5.2 (batch)OpenAI400K~533$0.88$7.00
27GPT-5.2 ProOpenAI400K~533$21.00$168.00
28GPT-5.2 Pro (batch)OpenAI400K~533$10.50$84.00
29GPT-5.2-CodexOpenAI400K~533$1.75$14.00
30GPT-5.3-CodexOpenAI400K~533$1.75$14.00
31GPT-5.4 MiniOpenAI400K~533$0.75$4.50
32GPT-5.4 Mini (batch)OpenAI400K~533$0.38$2.25
33GPT-5.4 NanoOpenAI400K~533$0.20$1.25
34GPT-5.4 Nano (batch)OpenAI400K~533$0.10$0.63
35Nova Lite 1.0Amazon300K~400$0.06$0.24
36Nova Pro 1.0Amazon300K~400$0.80$3.20
37GPT-5.4 Image 2OpenAI272K~363$8.00$15.00
38Aion 3.5aion-labs262K~350$3.00$6.00
39Aion 3.5 Miniaion-labs262K~350$0.70$1.40
40Devstral 2 2512Mistral AI262K~350$0.40$2.00
41Falcon-H1-Arabic 34B InstructTII262K~350FreeFree
42Falcon-H1-Arabic 7B InstructTII262K~350FreeFree
43Gemma 4 26B A4B Google262K~350$0.09$0.30
44Gemma 4 26B A4B (free)Google262K~350FreeFree
45Gemma 4 31BGoogle262K~350$0.09$0.34
46Gemma 4 31B (free)Google262K~350FreeFree
47Hy3Tencent262K~350$0.13$0.53
48Hy3 previewTencent262K~350$0.18$0.60
49KAT-Coder-Pro V2.5Kuaishou262K~350$0.74$2.96
50Kimi K2 0905Moonshot AI262K~350$0.60$2.50
51Kimi K2 ThinkingMoonshot AI262K~350$0.60$2.50
52Kimi K2.5Moonshot AI262K~350$0.45$2.25
53Kimi K2.6Moonshot AI262K~350$0.95$4.00
54Kimi K2.7 CodeMoonshot AI262K~350$0.66$3.30
55Laguna S 2.1 (free)poolside262K~350FreeFree
56Laguna XS 2.1poolside262K~350$0.06$0.12
57Laguna XS 2.1 (free)poolside262K~350FreeFree
58Ling 3.0 Flashinclusionai262K~350$0.02$0.06
59Ling 3.0 Flash Fininclusionai262K~350$0.06$0.18
60Ling 3.0 Flash Fin (free)inclusionai262K~350FreeFree
61Ling 3.0 Flash Sante (free)inclusionai262K~350FreeFree
62Ling 3.0 Flash VLinclusionai262K~350$0.06$0.18
63Ministral 3 14B 2512Mistral AI262K~350$0.20$0.20
64Ministral 3 8B 2512Mistral AI262K~350$0.15$0.15
65Ministral 3 8B 2512 (batch)Mistral AI262K~350$0.07$0.07
66Mistral Large 3 2512Mistral AI262K~350$0.50$1.50
67Mistral Large 3 2512 (batch)Mistral AI262K~350$0.25$0.75
68Mistral Medium 3.5Mistral AI262K~350$1.50$7.50
69Mistral Medium 3.5 (batch)Mistral AI262K~350$0.75$3.75
70Mistral Small 4Mistral AI262K~350$0.15$0.60
71Mistral Small 4 (batch)Mistral AI262K~350$0.07$0.30
72Nemotron 3 Nano 30B A3BNVIDIA262K~350$0.05$0.20
73Nemotron 3 SuperNVIDIA262K~350$0.08$0.45
74Nemotron 3 Super (free)NVIDIA262K~350FreeFree
75Nemotron 3 UltraNVIDIA262K~350$0.60$2.40
76Nemotron 3.5 LightningNVIDIA262K~350$0.08$0.20
77Paretounbiased262K~350$2.50$7.50
78Qwen3 235B A22B Instruct 2507Alibaba262K~350$0.09$0.35
79Qwen3 30B A3B Instruct 2507Alibaba262K~350$0.10$0.30
80Qwen3 Coder 30B A3B InstructAlibaba262K~350$0.07$0.28
81Qwen3 Coder 480B A35BAlibaba262K~350$0.30$1.00
82Qwen3 Coder NextAlibaba262K~350$0.12$0.80
83Qwen3 MaxAlibaba262K~350$0.78$3.90
84Qwen3 Max ThinkingAlibaba262K~350$0.78$3.90
85Qwen3 Next 80B A3B InstructAlibaba262K~350$0.10$1.10
86Qwen3 Next 80B A3B ThinkingAlibaba262K~350$0.15$1.20
87Qwen3 VL 235B A22B InstructAlibaba262K~350$0.21$1.90
88Qwen3 VL 30B A3B InstructAlibaba262K~350$0.15$0.60
89Qwen3 VL 30B A3B ThinkingAlibaba262K~350$0.20$2.40
90Qwen3 VL 8B InstructAlibaba262K~350$0.12$0.45
91Qwen3.5 397B A17BAlibaba262K~350$0.55$3.50
92Qwen3.5-122B-A10BAlibaba262K~350$0.26$2.08
93Qwen3.5-27BAlibaba262K~350$0.20$1.56
94Qwen3.5-35B-A3BAlibaba262K~350$0.31$1.25
95Qwen3.5-9BAlibaba262K~350$0.10$0.15
96Qwen3.6 27BAlibaba262K~350$0.32$2.70
97Qwen3.6 35B A3BAlibaba262K~350$0.15$1.00
98Qwen3.6 Max PreviewAlibaba262K~350$1.03$6.16
99Qwen3.8 27B (free)Alibaba262K~350FreeFree
100Sakana Namazusakana262K~350$0.95$4.00
101Seed 1.6ByteDance262K~350$0.25$2.00
102Seed 1.6 FlashByteDance262K~350$0.07$0.30
103Seed 2.1 TurboByteDance262K~350$0.50$2.50
104Seed-2.0-CodeByteDance262K~350$0.50$3.00
105Seed-2.0-LiteByteDance262K~350$0.25$2.00
106Seed-2.0-MiniByteDance262K~350$0.10$0.40
107Step 3.5 FlashStepFun262K~350$0.10$0.30
108Step 3.7 FlashStepFun262K~350$0.20$1.15
109Ternary Bonsai 2 27Bprism-ml262K~350$0.07$0.50
110Trinity Large Thinkingarcee-ai262K~350$0.25$0.80
111Mercury 2.5Inception260K~347$0.04$0.15
112Codestral 2508Mistral AI256K~341$0.30$0.90
113Codestral 2508 (batch)Mistral AI256K~341$0.15$0.45
114Command ACohere256K~341$2.50$10.00
115Grok Build 0.1xAI256K~341$1.00$2.00
116Mistral Small 3.2 24BMistral AI256K~341$0.09$0.25
117Nemotron 3 Nano Omni (free)NVIDIA256K~341FreeFree
118North Mini Code (free)Cohere256K~341FreeFree
119GLM 4.6Zhipu AI205K~273$0.43$1.75
120GLM 4.7Zhipu AI205K~273$0.60$2.20
121GLM 5Zhipu AI205K~273$0.60$1.92
122GLM 5.1Zhipu AI205K~273$0.96$3.03
123MiniMax M2MiniMax205K~273$0.30$1.20
124MiniMax M2.1MiniMax205K~273$0.30$1.20
125MiniMax M2.5MiniMax205K~273$0.27$1.08
126MiniMax M2.7MiniMax205K~273$0.30$1.20
127GLM 5 TurboZhipu AI203K~270$1.20$4.00
128GLM 5V TurboZhipu AI203K~270$1.20$4.00
129Claude 3 HaikuAnthropic200K~267$0.25$1.25
130Claude Haiku 4.5Anthropic200K~267$1.00$5.00
131Claude Haiku 4.5 (batch)Anthropic200K~267$0.50$2.50
132Claude Haiku Latest~anthropic200K~267$1.00$5.00
133Claude Opus 4.1Anthropic200K~267$15.00$75.00
134Claude Opus 4.1 (batch)Anthropic200K~267$7.50$37.50
135Claude Opus 4.5Anthropic200K~267$5.00$25.00
136Claude Opus 4.5 (batch)Anthropic200K~267$2.50$12.50
137Claude Sonnet 4Anthropic200K~267$3.00$15.00
138Composer 2Cursor200K~267$0.50$2.50
139Composer 2 FastCursor200K~267$1.50$7.50
140GLM 4.7 FlashZhipu AI200K~267$0.06$0.40
141o1OpenAI200K~267$15.00$60.00
142o1-proOpenAI200K~267$150.00$600.00
143o3OpenAI200K~267$2.00$8.00
144o3 (batch)OpenAI200K~267$1.00$4.00
145o3 MiniOpenAI200K~267$1.10$4.40
146o3 Mini (batch)OpenAI200K~267$0.55$2.20
147o3 Mini HighOpenAI200K~267$1.10$4.40
148o3 ProOpenAI200K~267$20.00$80.00
149o4 MiniOpenAI200K~267$1.10$4.40
150o4 Mini (batch)OpenAI200K~267$0.55$2.20
151o4 Mini HighOpenAI200K~267$1.10$4.40
152Sonar ProPerplexity200K~267$3.00$15.00
153Sonar Pro SearchPerplexity200K~267$3.00$15.00

128K Tokens

(79 models)

The current standard for frontier models. Sufficient for most production use cases, including long conversations and medium-length documents.

#ModelProviderContext WindowPagesInput / 1MOutput / 1M
1Command A+Cohere192K~256$0.30$1.50
2DeepSeek V3DeepSeek164K~218$0.32$0.89
3DeepSeek V3 0324DeepSeek164K~218$0.25$1.00
4DeepSeek V3.1DeepSeek164K~218$0.25$0.95
5DeepSeek V3.1 TerminusDeepSeek164K~218$0.27$1.00
6DeepSeek V3.2DeepSeek164K~218$0.27$0.40
7DeepSeek V3.2 ExpDeepSeek164K~218$0.27$0.41
8Llama Guard 4 12BMeta164K~218$0.18$0.18
9R1 0528DeepSeek164K~218$0.50$2.15
10Aion-2.0aion-labs131K~175$0.80$1.60
11Aion-3.0aion-labs131K~175$3.00$6.00
12Aion-3.0-Miniaion-labs131K~175$0.70$1.40
13Falcon-H1-Arabic 3B InstructTII131K~175FreeFree
14Gemma 3 12BGoogle131K~175$0.05$0.15
15Gemma 3 27BGoogle131K~175$0.08$0.45
16Gemma 3 4BGoogle131K~175$0.05$0.10
17GLM 4.5Zhipu AI131K~175$0.60$2.20
18GLM 4.5 AirZhipu AI131K~175$0.13$0.85
19GLM 4.6VZhipu AI131K~175$0.30$0.90
20gpt-oss-120bOpenAI131K~175$0.15$0.60
21gpt-oss-120b (batch)OpenAI131K~175$0.03$0.14
22gpt-oss-20bOpenAI131K~175$0.02$0.09
23gpt-oss-20b (batch)OpenAI131K~175$0.02$0.11
24gpt-oss-safeguard-20bOpenAI131K~175$0.07$0.30
25Granite 4.2 8BIBM131K~175$0.06$0.25
26Hunyuan A13B InstructTencent131K~175$0.14$0.57
27Kimi K2 0711Moonshot AI131K~175$0.57$2.30
28Llama 3.1 70B InstructMeta131K~175$0.40$0.40
29Llama 3.1 8B InstructMeta131K~175$0.05$0.08
30Llama 3.2 3B InstructMeta131K~175$0.05$0.33
31Llama 3.3 70B InstructMeta131K~175$0.10$0.32
32Ministral 3 3B 2512Mistral AI131K~175$0.10$0.10
33Mistral Large 2407Mistral AI131K~175$2.00$6.00
34Mistral Medium 3Mistral AI131K~175$0.40$2.00
35Mistral Medium 3.1Mistral AI131K~175$0.40$2.00
36Mistral Medium 3.1 (batch)Mistral AI131K~175$0.20$1.00
37Mistral NemoMistral AI131K~175$0.02$0.03
38Muse Glimmer 30Bmeta131K~175$0.30$1.20
39Nano Banana 2 (Gemini 3.1 Flash Image)Google131K~175$0.50$3.00
40Nano Banana Pro (Gemini 3 Pro Image)Google131K~175$2.00$12.00
41Nemotron 3.5 Content SafetyNVIDIA131K~175$0.20$0.20
42Qwen3 14BAlibaba131K~175$0.12$0.24
43Qwen3 235B A22BAlibaba131K~175$0.45$1.82
44Qwen3 235B A22B Thinking 2507Alibaba131K~175$0.23$2.30
45Qwen3 30B A3BAlibaba131K~175$0.12$0.50
46Qwen3 32BAlibaba131K~175$0.08$0.28
47Qwen3 8BAlibaba131K~175$0.12$0.45
48Qwen3 VL 235B A22B ThinkingAlibaba131K~175$0.40$4.00
49Qwen3 VL 32B InstructAlibaba131K~175$0.10$0.42
50Qwen3 VL 8B ThinkingAlibaba131K~175$0.18$2.10
51Solar Pro 3Upstage131K~175$0.15$0.60
52Granite 4.0 MicroIBM131K~175$0.02$0.11
53Command R (08-2024)Cohere128K~171$0.15$0.60
54Command R+ (08-2024)Cohere128K~171$2.50$10.00
55Command R7B (12-2024)Cohere128K~171$0.04$0.15
56GPT AudioOpenAI128K~171$2.50$10.00
57GPT Audio MiniOpenAI128K~171$0.60$2.40
58GPT-4 TurboOpenAI128K~171$10.00$30.00
59GPT-4 Turbo (batch)OpenAI128K~171$5.00$15.00
60GPT-4oOpenAI128K~171$2.50$10.00
61GPT-4o (2024-05-13)OpenAI128K~171$5.00$15.00
62GPT-4o (2024-08-06)OpenAI128K~171$2.50$10.00
63GPT-4o (2024-11-20)OpenAI128K~171$2.50$10.00
64GPT-4o (batch)OpenAI128K~171$1.25$5.00
65GPT-4o-miniOpenAI128K~171$0.15$0.60
66GPT-4o-mini (2024-07-18)OpenAI128K~171$0.15$0.60
67GPT-4o-mini (batch)OpenAI128K~171$0.07$0.30
68GPT-5.2 ChatOpenAI128K~171$1.75$14.00
69Mercury 2Inception128K~171$0.25$0.75
70Mistral LargeMistral AI128K~171$2.00$6.00
71Mistral Small 3.1 24BMistral AI128K~171$0.35$0.55
72Nemotron 3.5 Content Safety (free)NVIDIA128K~171FreeFree
73Nova Micro 1.0Amazon128K~171$0.04$0.14
74Qwen2.5 VL 72B InstructAlibaba128K~171$0.80$1.00
75Schematron V2 Smallinference-net128K~171$0.05$0.23
76Schematron V2 Turboinference-net128K~171$0.03$0.15
77Sonar Deep ResearchPerplexity128K~171$2.00$8.00
78Sonar Reasoning ProPerplexity128K~171$2.00$8.00
79UI-TARS 7B ByteDance128K~171$0.10$0.20

32K–64K Tokens

(27 models)

Moderate context windows suitable for shorter documents, code files, and focused conversations.

#ModelProviderContext WindowPagesInput / 1MOutput / 1M
1SonarPerplexity127K~169$1.00$1.00
2ERNIE 4.5 VL 424B A47B Baidu123K~164$0.42$1.25
3Qwen3 30B A3B Thinking 2507Alibaba82K~109$0.20$2.40
4GLM 4.5VZhipu AI66K~87$0.60$1.80
5LFM2.5-2.6B (free)Liquid AI66K~87FreeFree
6MiniMax M2-herMiniMax66K~87$0.30$1.20
7Mixtral 8x22B InstructMistral AI66K~87$2.00$6.00
8Nano Banana 2 (Gemini 3.1 Flash Image Preview)Google66K~87$0.50$3.00
9Nano Banana 2 Lite (Gemini 3.1 Flash Lite Image)Google66K~87$0.25$1.50
10Nano Banana Pro (Gemini 3 Pro Image Preview)Google66K~87$2.00$12.00
11Reka Flash 3rekaai66K~87$0.10$0.20
12WizardLM-2 8x22BMicrosoft66K~87$0.62$0.62
13R1DeepSeek64K~85$0.70$2.50
14Llama 3.2 1B InstructMeta60K~80$0.03$0.20
15Falcon Arabic 7B InstructTII33K~44FreeFree
16Falcon Mamba 7B InstructTII33K~44FreeFree
17Falcon3 10B InstructTII33K~44FreeFree
18Falcon3 7B InstructTII33K~44FreeFree
19GLM 5.2 (free)Zhipu AI33K~44FreeFree
20Mistral Small 3Mistral AI33K~44$0.05$0.08
21Nano Banana (Gemini 2.5 Flash Image)Google33K~44$0.30$2.50
22Perceptron Mk1perceptron33K~44$0.15$1.50
23Qwen2.5 72B InstructAlibaba33K~44$0.36$0.40
24Qwen2.5 7B InstructAlibaba33K~44$0.10$0.20
25Qwen2.5 Coder 32B InstructAlibaba33K~44$0.66$1.00
26SabaMistral AI33K~44$0.20$0.60
27Voxtral Small 24B 2507Mistral AI33K~44$0.10$0.30

What Is a Context Window and Why Does It Matter?

Context window = working memory

A model's context window is the total number of tokens (roughly words) it can process in a single request. This includes both your input prompt and the model's output. A model with a 128K context window can process about 170 pages of text at once, while a 1M-token model can handle roughly 1,300 pages -- enough for entire books or large codebases.

Larger context enables new use cases

With small context windows (under 32K), you must chunk documents and use retrieval-augmented generation (RAG). Large context models eliminate this complexity for many workloads: analyzing full legal contracts, reviewing entire repositories, summarizing research paper collections, or maintaining very long conversations with full history retained.

Context size vs. effective recall

Not all context is created equal. Some models perform well on "needle in a haystack" tests at their full context length, while others degrade on information retrieval when prompts get very long. The advertised context window is the maximum, but effective performance may vary. Check our leaderboard for quality scores that account for real-world performance.

Cost implications of large context

Using a large context window means sending more tokens per request, which increases cost. For example, filling a 1M-token context at $3/1M input tokens costs $3 per request. For cost-sensitive workloads, consider whether RAG with a smaller context model might be more efficient than filling a large context window end to end.

Explore More

Dive deeper into model capabilities, compare context windows side by side, or see overall rankings across all dimensions.

Frequently Asked Questions

A large context window model can process long inputs - from 128K tokens (about 100 pages) up to 2M tokens (about 1,500 pages). This allows analyzing entire codebases, long documents, or extensive conversation histories in a single request.

Large context models excel when you need the AI to reason across an entire document at once. RAG (Retrieval-Augmented Generation) is better for searching across massive document collections. Large context is simpler to implement but costs more per request.

No. While many models advertise 1M+ token context windows, their effective recall can vary significantly. Some models lose accuracy when important information is buried in the middle of very long contexts. Our rankings account for practical performance, not just advertised limits.

Large Context AI Models - 1M+ Token (2026) | LM Market Cap