Skip to content

Large Context Window AI Models

Compare 421 AI models with context windows of 32K tokens or more. The largest models support 2M tokens -- enough to process entire codebases, books, or hundreds of documents in a single prompt. Data updated hourly.

166
1M+ Tokens
149
200K+ Tokens
79
128K Tokens
27
32K–64K Tokens
2M
Largest Context
421
Models (32K+)
545K
Average Context

1M+ Tokens

(166 models)

The largest context windows available. These models can process entire codebases, full books, or hundreds of documents in a single request.

#ModelProviderContext WindowPagesInput / 1MOutput / 1M
1Grok 4.20xAI2M~3K$1.25$2.50
2Grok 4.20 Multi-AgentxAI2M~3K$1.25$2.50
3DeepSeek V4 Flash 0731DeepSeek1.3M~2K$0.04$0.64
4DeepSeek V4 Flash Latest~deepseek1.3M~2K$0.03$0.80
5GLM 5.3Zhipu AI1.3M~2K$0.56$1.76
6GLM 5.3 FlashZhipu AI1.3M~2K$0.15$0.50
7GLM Flash Latest~z-ai1.3M~2K$0.07$0.25
8GLM Latest~z-ai1.3M~2K$0.56$1.76
9Llama 4 ScoutMeta1.3M~2K$0.10$0.30
10GPT Astra Latest~openai1.1M~1K$10.00$50.00
11GPT Luna Latest~openai1.1M~1K$0.10$0.50
12GPT Sol Latest~openai1.1M~1K$2.00$10.00
13GPT Terra Latest~openai1.1M~1K$2.00$12.00
14GPT-5.4OpenAI1.1M~1K$2.50$15.00
15GPT-5.4 (batch)OpenAI1.1M~1K$1.25$7.50
16GPT-5.4 ProOpenAI1.1M~1K$30.00$180.00
17GPT-5.4 Pro (batch)OpenAI1.1M~1K$15.00$90.00
18GPT-5.5OpenAI1.1M~1K$5.00$30.00
19GPT-5.5 (batch)OpenAI1.1M~1K$2.50$15.00
20GPT-5.5 ProOpenAI1.1M~1K$30.00$180.00
21GPT-5.5 Pro (batch)OpenAI1.1M~1K$15.00$90.00
22GPT-5.6 LunaOpenAI1.1M~1K$0.20$1.20
23GPT-5.6 Luna (batch)OpenAI1.1M~1K$0.10$0.60
24GPT-5.6 Luna ProOpenAI1.1M~1K$0.20$1.20
25GPT-5.6 Luna Pro (batch)OpenAI1.1M~1K$0.10$0.60
26GPT-5.6 SolOpenAI1.1M~1K$2.00$10.00
27GPT-5.6 Sol (batch)OpenAI1.1M~1K$1.00$5.00
28GPT-5.6 Sol ProOpenAI1.1M~1K$2.00$10.00
29GPT-5.6 Sol Pro (batch)OpenAI1.1M~1K$1.00$5.00
30GPT-5.6 TerraOpenAI1.1M~1K$2.00$12.00
31GPT-5.6 Terra (batch)OpenAI1.1M~1K$1.00$6.00
32GPT-5.6 Terra ProOpenAI1.1M~1K$2.00$12.00
33GPT-5.6 Terra Pro (batch)OpenAI1.1M~1K$1.00$6.00
34GPT-6 AstraOpenAI1.1M~1K$10.00$50.00
35GPT-6 Astra (batch)OpenAI1.1M~1K$5.00$25.00
36GPT-6 Astra ProOpenAI1.1M~1K$10.00$50.00
37GPT-6 Astra Pro (batch)OpenAI1.1M~1K$5.00$25.00
38GPT-6 LunaOpenAI1.1M~1K$0.10$0.50
39GPT-6 Luna (batch)OpenAI1.1M~1K$0.05$0.25
40GPT-6 Luna ProOpenAI1.1M~1K$0.10$0.50
41GPT-6 Luna Pro (batch)OpenAI1.1M~1K$0.05$0.25
42GPT-6 SolOpenAI1.1M~1K$2.00$10.00
43GPT-6 Sol (batch)OpenAI1.1M~1K$1.00$5.00
44GPT-6 Sol ProOpenAI1.1M~1K$2.00$10.00
45GPT-6 Sol Pro (batch)OpenAI1.1M~1K$1.00$5.00
46MiMo-V2.5Xiaomi1.1M~1K$0.14$0.28
47MiMo-V2.5-ProXiaomi1.1M~1K$0.43$0.87
48LongCat 2.0Meituan1.0M~1K$0.30$1.20
49DeepSeek Flash Latest~deepseek1.0M~1K$0.10$0.50
50DeepSeek Pro Latest~deepseek1.0M~1K$0.40$1.20
51DeepSeek V4 Flash 0423DeepSeek1.0M~1K$0.05$0.10
52DeepSeek V4 Flash Vision ExpDeepSeek1.0M~1K$0.22$0.66
53DeepSeek V4 Pro 0423DeepSeek1.0M~1K$0.86$1.72
54DeepSeek V4 Pro 0813DeepSeek1.0M~1K$0.66$1.98
55DeepSeek V4.1 FlashDeepSeek1.0M~1K$0.10$0.50
56DeepSeek V4.1 Flash (batch)DeepSeek1.0M~1K$0.11$0.34
57Gemini 2.5 FlashGoogle1.0M~1K$0.30$2.50
58Gemini 2.5 Flash (batch)Google1.0M~1K$0.15$1.25
59Gemini 2.5 Flash LiteGoogle1.0M~1K$0.10$0.40
60Gemini 2.5 Flash Lite (batch)Google1.0M~1K$0.05$0.20
61Gemini 2.5 ProGoogle1.0M~1K$1.25$10.00
62Gemini 2.5 Pro (batch)Google1.0M~1K$0.63$5.00
63Gemini 2.5 Pro Preview 06-05Google1.0M~1K$1.25$10.00
64Gemini 3 Flash PreviewGoogle1.0M~1K$0.50$3.00
65Gemini 3 Flash Preview (batch)Google1.0M~1K$0.25$1.50
66Gemini 3.1 Flash LiteGoogle1.0M~1K$0.25$1.50
67Gemini 3.1 Flash Lite (batch)Google1.0M~1K$0.13$0.75
68Gemini 3.1 Flash Lite PreviewGoogle1.0M~1K$0.25$1.50
69Gemini 3.1 Pro PreviewGoogle1.0M~1K$2.00$12.00
70Gemini 3.1 Pro Preview (batch)Google1.0M~1K$1.00$6.00
71Gemini 3.1 Pro Preview Custom ToolsGoogle1.0M~1K$2.00$12.00
72Gemini 3.5 FlashGoogle1.0M~1K$1.50$9.00
73Gemini 3.5 Flash (batch)Google1.0M~1K$0.75$4.50
74Gemini 3.5 Flash LiteGoogle1.0M~1K$0.30$2.50
75Gemini 3.5 Flash Lite (batch)Google1.0M~1K$0.15$1.25
76Gemini 3.6 FlashGoogle1.0M~1K$0.75$3.75
77Gemini 3.6 Flash (batch)Google1.0M~1K$0.38$1.88
78Gemini 3.7 FlashGoogle1.0M~1K$0.75$3.75
79Gemini 3.7 Flash (batch)Google1.0M~1K$0.38$1.88
80Gemini 3.8 FlashGoogle1.0M~1K$0.75$3.75
81Gemini 3.8 Flash (batch)Google1.0M~1K$0.38$1.88
82Gemini Flash Latest~google1.0M~1K$0.75$3.75
83Gemini Pro Latest~google1.0M~1K$2.00$12.00
84GLM 5.2Zhipu AI1.0M~1K$0.65$2.04
85GLM 5.3 (batch)Zhipu AI1.0M~1K$0.72$2.40
86GLM 5.3 Flash (batch)Zhipu AI1.0M~1K$0.06$0.20
87GLM 5.3 FlashXZhipu AI1.0M~1K$0.37$1.25
88Hy4 previewTencent1.0M~1K$0.83$2.50
89Inklingthinkingmachines1.0M~1K$1.00$4.05
90Inkling (free)thinkingmachines1.0M~1KFreeFree
91Inkling Smallthinkingmachines1.0M~1K$0.45$1.20
92Inkling Small (free)thinkingmachines1.0M~1KFreeFree
93Kimi K3Moonshot AI1.0M~1K$3.00$15.00
94Kimi K3 (batch)Moonshot AI1.0M~1K$2.28$11.40
95Kimi Latest~moonshotai1.0M~1K$1.50$7.50
96Laguna S 2.1poolside1.0M~1K$0.09$0.18
97Llama 4 MaverickMeta1.0M~1K$0.19$0.65
98Lyria 3 Clip PreviewGoogle1.0M~1KFreeFree
99Lyria 3 Pro PreviewGoogle1.0M~1KFreeFree
100MiMo-V2.6-FlashXiaomi1.0M~1K$0.14$0.28
101MiMo-V2.6-ProXiaomi1.0M~1K$0.43$0.87
102MiMo-V2.6-Pro-UltraSpeedXiaomi1.0M~1K$4.35$8.70
103MiniMax M3MiniMax1.0M~1K$0.30$1.20
104Muse Spark 1.1meta1.0M~1K$1.25$4.25
105Muse Spark 1.2meta1.0M~1K$1.25$4.25
106Muse Spark 1.2 Contributormeta1.0M~1K$0.10$0.20
107Muse Spark 1.3meta1.0M~1K$1.25$4.25
108Muse Spark 1.3 Contributormeta1.0M~1K$0.10$0.20
109Qwen3.8 2.4T A95BAlibaba1.0M~1K$2.00$6.00
110GPT-4.1OpenAI1.0M~1K$2.00$8.00
111GPT-4.1 (batch)OpenAI1.0M~1K$1.00$4.00
112GPT-4.1 MiniOpenAI1.0M~1K$0.40$1.60
113GPT-4.1 Mini (batch)OpenAI1.0M~1K$0.20$0.80
114GPT-4.1 NanoOpenAI1.0M~1K$0.10$0.40
115GPT-4.1 Nano (batch)OpenAI1.0M~1K$0.05$0.20
116Palmyra X5Writer1.0M~1K$0.60$6.00
117MiniMax-01MiniMax1.0M~1K$0.20$1.10
118Claude Fable 5Anthropic1M~1K$10.00$50.00
119Claude Fable 5 (batch)Anthropic1M~1K$5.00$25.00
120Claude Fable 5.1Anthropic1M~1K$10.00$50.00
121Claude Fable 5.1 (batch)Anthropic1M~1K$5.00$25.00
122Claude Fable Latest~anthropic1M~1K$10.00$50.00
123Claude Opus 4.6Anthropic1M~1K$5.00$25.00
124Claude Opus 4.6 (batch)Anthropic1M~1K$2.50$12.50
125Claude Opus 4.7Anthropic1M~1K$5.00$25.00
126Claude Opus 4.7 (batch)Anthropic1M~1K$2.50$12.50
127Claude Opus 4.8Anthropic1M~1K$5.00$25.00
128Claude Opus 4.8 (batch)Anthropic1M~1K$2.50$12.50
129Claude Opus 5Anthropic1M~1K$5.00$25.00
130Claude Opus 5 (batch)Anthropic1M~1K$2.50$12.50
131Claude Opus 5.5Anthropic1M~1K$4.00$20.00
132Claude Opus 5.5 (batch)Anthropic1M~1K$2.00$10.00
133Claude Opus Latest~anthropic1M~1K$4.00$20.00
134Claude Sonnet 4.5Anthropic1M~1K$3.00$15.00
135Claude Sonnet 4.5 (batch)Anthropic1M~1K$1.50$7.50
136Claude Sonnet 4.6Anthropic1M~1K$3.00$15.00
137Claude Sonnet 4.6 (batch)Anthropic1M~1K$1.50$7.50
138Claude Sonnet 5Anthropic1M~1K$2.00$10.00
139Claude Sonnet 5 (batch)Anthropic1M~1K$1.00$5.00
140Claude Sonnet Latest~anthropic1M~1K$2.00$10.00
141Fugu Maxsakana1M~1K$2.00$6.00
142Fugu Ultrasakana1M~1K$5.00$30.00
143Fugu Ultra v2sakana1M~1K$5.00$30.00
144Grok 4.3xAI1M~1K$1.25$2.50
145Grok 4.3 (batch)xAI1M~1K$1.00$2.00
146MiniMax M1MiniMax1M~1K$0.40$2.20
147Nemotron 3 Ultra (free)NVIDIA1M~1KFreeFree
148Nemotron 3.5 Lightning (free)NVIDIA1M~1KFreeFree
149Nova 2 LiteAmazon1M~1K$0.30$2.50
150Nova Premier 1.0Amazon1M~1K$2.50$12.50
151Qwen Plus 0728Alibaba1M~1K$0.26$0.78
152Qwen-PlusAlibaba1M~1K$0.26$0.78
153Qwen3 Coder FlashAlibaba1M~1K$0.20$0.97
154Qwen3 Coder PlusAlibaba1M~1K$0.65$3.25
155Qwen3.5 Plus 2026-02-15Alibaba1M~1K$0.26$1.56
156Qwen3.5 Plus 2026-04-20Alibaba1M~1K$0.30$1.80
157Qwen3.5-FlashAlibaba1M~1K$0.07$0.26
158Qwen3.6 FlashAlibaba1M~1K$0.19$1.13
159Qwen3.6 PlusAlibaba1M~1K$0.33$1.95
160Qwen3.7 FlashAlibaba1M~1K$0.03$0.13
161Qwen3.7 MaxAlibaba1M~1K$1.48$4.42
162Qwen3.7 PlusAlibaba1M~1K$0.32$1.28
163Qwen3.8 27BAlibaba1M~1K$0.42$3.00
164Qwen3.8 FlashAlibaba1M~1K$0.15$0.47
165Qwen3.8 Max (0902)Alibaba1M~1K$2.00$6.00
166Qwen3.8 Omni FlashAlibaba1M~1K$0.15$0.47

200K+ Tokens

(149 models)

Extended context models ideal for long documents, legal contracts, research papers, and multi-file code analysis.

#ModelProviderContext WindowPagesInput / 1MOutput / 1M
1Solar Pro 4Upstage524K~699$0.09$0.36
2Dots3-Note Preview (free)dots-studio512K~683FreeFree
3Grok 4.5xAI500K~667$2.00$6.00
4Grok 4.6xAI500K~667$2.00$6.00
5Grok 4.7xAI500K~667$1.60$4.80
6Grok Latest~x-ai500K~667$1.60$4.80
7GPT Chat LatestOpenAI400K~533$5.00$30.00
8GPT Mini Latest~openai400K~533$0.75$4.50
9GPT-5OpenAI400K~533$1.25$10.00
10GPT-5 (batch)OpenAI400K~533$0.63$5.00
11GPT-5 ImageOpenAI400K~533$10.00$10.00
12GPT-5 Image MiniOpenAI400K~533$2.50$2.00
13GPT-5 MiniOpenAI400K~533$0.25$2.00
14GPT-5 Mini (batch)OpenAI400K~533$0.13$1.00
15GPT-5 NanoOpenAI400K~533$0.05$0.40
16GPT-5 Nano (batch)OpenAI400K~533$0.02$0.20
17GPT-5 ProOpenAI400K~533$15.00$120.00
18GPT-5 Pro (batch)OpenAI400K~533$7.50$60.00
19GPT-5.1OpenAI400K~533$1.25$10.00
20GPT-5.1 (batch)OpenAI400K~533$0.63$5.00
21GPT-5.1-CodexOpenAI400K~533$1.25$10.00
22GPT-5.1-Codex-MaxOpenAI400K~533$1.25$10.00
23GPT-5.1-Codex-MiniOpenAI400K~533$0.25$2.00
24GPT-5.2OpenAI400K~533$1.75$14.00
25GPT-5.2 (batch)OpenAI400K~533$0.88$7.00
26GPT-5.2 ProOpenAI400K~533$21.00$168.00
27GPT-5.2 Pro (batch)OpenAI400K~533$10.50$84.00
28GPT-5.2-CodexOpenAI400K~533$1.75$14.00
29GPT-5.3-CodexOpenAI400K~533$1.75$14.00
30GPT-5.4 MiniOpenAI400K~533$0.75$4.50
31GPT-5.4 Mini (batch)OpenAI400K~533$0.38$2.25
32GPT-5.4 NanoOpenAI400K~533$0.20$1.25
33GPT-5.4 Nano (batch)OpenAI400K~533$0.10$0.63
34Nova Lite 1.0Amazon300K~400$0.06$0.24
35Nova Pro 1.0Amazon300K~400$0.80$3.20
36GPT-5.4 Image 2OpenAI272K~363$8.00$15.00
37Devstral 2 2512Mistral AI262K~350$0.40$2.00
38Falcon-H1-Arabic 34B InstructTII262K~350FreeFree
39Falcon-H1-Arabic 7B InstructTII262K~350FreeFree
40Gemma 4 26B A4B Google262K~350$0.09$0.30
41Gemma 4 26B A4B (free)Google262K~350FreeFree
42Gemma 4 31BGoogle262K~350$0.09$0.34
43Gemma 4 31B (free)Google262K~350FreeFree
44Hy3Tencent262K~350$0.08$0.33
45Hy3 previewTencent262K~350$0.18$0.60
46KAT-Coder-Pro V2.5Kuaishou262K~350$0.74$2.96
47Kimi K2 0905Moonshot AI262K~350$0.60$2.50
48Kimi K2 ThinkingMoonshot AI262K~350$0.60$2.50
49Kimi K2.5Moonshot AI262K~350$0.45$2.25
50Kimi K2.6Moonshot AI262K~350$0.95$4.00
51Kimi K2.7 CodeMoonshot AI262K~350$0.71$3.30
52Laguna S 2.1 (free)poolside262K~350FreeFree
53Laguna XS 2.1poolside262K~350$0.06$0.12
54Laguna XS 2.1 (free)poolside262K~350FreeFree
55Ling 3.0 Flashinclusionai262K~350$0.02$0.06
56Ling 3.0 Flash Fininclusionai262K~350$0.06$0.18
57Ling 3.0 Flash Fin (free)inclusionai262K~350FreeFree
58Ling 3.0 Flash Sante (free)inclusionai262K~350FreeFree
59Ling 3.0 Flash VL (free)inclusionai262K~350FreeFree
60Ministral 3 14B 2512Mistral AI262K~350$0.20$0.20
61Ministral 3 8B 2512Mistral AI262K~350$0.15$0.15
62Ministral 3 8B 2512 (batch)Mistral AI262K~350$0.07$0.07
63Mistral Large 3 2512 (batch)Mistral AI262K~350$0.25$0.75
64Mistral Medium 3.5Mistral AI262K~350$1.50$7.50
65Mistral Medium 3.5 (batch)Mistral AI262K~350$0.75$3.75
66Mistral Small 4Mistral AI262K~350$0.15$0.60
67Mistral Small 4 (batch)Mistral AI262K~350$0.07$0.30
68Nemotron 3 Nano 30B A3BNVIDIA262K~350$0.05$0.20
69Nemotron 3 SuperNVIDIA262K~350$0.08$0.45
70Nemotron 3 Super (free)NVIDIA262K~350FreeFree
71Nemotron 3 UltraNVIDIA262K~350$0.60$2.40
72Nemotron 3.5 LightningNVIDIA262K~350$0.07$0.20
73Paretounbiased262K~350$2.50$7.50
74Qwen3 235B A22B Instruct 2507Alibaba262K~350$0.09$0.35
75Qwen3 30B A3B Instruct 2507Alibaba262K~350$0.05$0.19
76Qwen3 Coder 30B A3B InstructAlibaba262K~350$0.07$0.28
77Qwen3 Coder 480B A35BAlibaba262K~350$0.30$1.00
78Qwen3 Coder NextAlibaba262K~350$0.12$0.80
79Qwen3 MaxAlibaba262K~350$0.78$3.90
80Qwen3 Max ThinkingAlibaba262K~350$0.78$3.90
81Qwen3 Next 80B A3B InstructAlibaba262K~350$0.09$1.10
82Qwen3 Next 80B A3B ThinkingAlibaba262K~350$0.15$1.20
83Qwen3 VL 235B A22B InstructAlibaba262K~350$0.21$1.90
84Qwen3 VL 30B A3B InstructAlibaba262K~350$0.13$0.52
85Qwen3 VL 30B A3B ThinkingAlibaba262K~350$0.20$2.40
86Qwen3 VL 8B InstructAlibaba262K~350$0.12$0.45
87Qwen3.5 397B A17BAlibaba262K~350$0.55$3.50
88Qwen3.5-122B-A10BAlibaba262K~350$0.26$2.08
89Qwen3.5-27BAlibaba262K~350$0.20$1.56
90Qwen3.5-35B-A3BAlibaba262K~350$0.31$1.25
91Qwen3.5-9BAlibaba262K~350$0.10$0.15
92Qwen3.6 27BAlibaba262K~350$0.32$2.70
93Qwen3.6 35B A3BAlibaba262K~350$0.15$1.00
94Qwen3.6 Max PreviewAlibaba262K~350$1.03$6.16
95Qwen3.8 27B (free)Alibaba262K~350FreeFree
96Sakana Namazusakana262K~350$0.95$4.00
97Seed 1.6ByteDance262K~350$0.25$2.00
98Seed 1.6 FlashByteDance262K~350$0.07$0.30
99Seed 2.1 TurboByteDance262K~350$0.50$2.50
100Seed-2.0-CodeByteDance262K~350$0.50$3.00
101Seed-2.0-LiteByteDance262K~350$0.25$2.00
102Seed-2.0-MiniByteDance262K~350$0.10$0.40
103Step 3.5 FlashStepFun262K~350$0.10$0.30
104Step 3.7 FlashStepFun262K~350$0.20$1.15
105Ternary Bonsai 2 27Bprism-ml262K~350$0.07$0.50
106Trinity Large Thinkingarcee-ai262K~350$0.25$0.80
107Mercury 2.5Inception260K~347$0.04$0.15
108Codestral 2508Mistral AI256K~341$0.30$0.90
109Codestral 2508 (batch)Mistral AI256K~341$0.15$0.45
110Command ACohere256K~341$2.50$10.00
111Grok Build 0.1xAI256K~341$1.00$2.00
112Mistral Small 3.2 24BMistral AI256K~341$0.09$0.25
113Nemotron 3 Nano Omni (free)NVIDIA256K~341FreeFree
114North Mini Code (free)Cohere256K~341FreeFree
115GLM 4.6Zhipu AI205K~273$0.43$1.75
116GLM 4.7Zhipu AI205K~273$0.40$1.75
117GLM 5Zhipu AI205K~273$0.60$1.92
118GLM 5.1Zhipu AI205K~273$0.97$3.04
119MiniMax M2MiniMax205K~273$0.26$1.02
120MiniMax M2.1MiniMax205K~273$0.30$1.20
121MiniMax M2.5MiniMax205K~273$0.27$1.08
122MiniMax M2.7MiniMax205K~273$0.30$1.20
123GLM 5 TurboZhipu AI203K~270$1.20$4.00
124GLM 5V TurboZhipu AI203K~270$1.20$4.00
125Claude 3 HaikuAnthropic200K~267$0.25$1.25
126Claude Haiku 4.5Anthropic200K~267$1.00$5.00
127Claude Haiku 4.5 (batch)Anthropic200K~267$0.50$2.50
128Claude Haiku Latest~anthropic200K~267$1.00$5.00
129Claude Opus 4.1Anthropic200K~267$15.00$75.00
130Claude Opus 4.1 (batch)Anthropic200K~267$7.50$37.50
131Claude Opus 4.5Anthropic200K~267$5.00$25.00
132Claude Opus 4.5 (batch)Anthropic200K~267$2.50$12.50
133Claude Sonnet 4Anthropic200K~267$3.00$15.00
134Composer 2Cursor200K~267$0.50$2.50
135Composer 2 FastCursor200K~267$1.50$7.50
136GLM 4.7 FlashZhipu AI200K~267$0.06$0.40
137o1OpenAI200K~267$15.00$60.00
138o1-proOpenAI200K~267$150.00$600.00
139o3OpenAI200K~267$2.00$8.00
140o3 (batch)OpenAI200K~267$1.00$4.00
141o3 MiniOpenAI200K~267$1.10$4.40
142o3 Mini (batch)OpenAI200K~267$0.55$2.20
143o3 Mini HighOpenAI200K~267$1.10$4.40
144o3 ProOpenAI200K~267$20.00$80.00
145o4 MiniOpenAI200K~267$1.10$4.40
146o4 Mini (batch)OpenAI200K~267$0.55$2.20
147o4 Mini HighOpenAI200K~267$1.10$4.40
148Sonar ProPerplexity200K~267$3.00$15.00
149Sonar Pro SearchPerplexity200K~267$3.00$15.00

128K Tokens

(79 models)

The current standard for frontier models. Sufficient for most production use cases, including long conversations and medium-length documents.

#ModelProviderContext WindowPagesInput / 1MOutput / 1M
1Command A+Cohere192K~256$0.30$1.50
2DeepSeek V3DeepSeek164K~218$0.32$0.89
3DeepSeek V3 0324DeepSeek164K~218$0.25$1.00
4DeepSeek V3.1DeepSeek164K~218$0.25$0.95
5DeepSeek V3.1 TerminusDeepSeek164K~218$0.27$1.00
6DeepSeek V3.2DeepSeek164K~218$0.27$0.40
7DeepSeek V3.2 ExpDeepSeek164K~218$0.27$0.41
8Llama Guard 4 12BMeta164K~218$0.18$0.18
9R1 0528DeepSeek164K~218$0.50$2.15
10Aion-2.0aion-labs131K~175$0.80$1.60
11Aion-3.0aion-labs131K~175$3.00$6.00
12Aion-3.0-Miniaion-labs131K~175$0.70$1.40
13Falcon-H1-Arabic 3B InstructTII131K~175FreeFree
14Gemma 3 12BGoogle131K~175$0.05$0.15
15Gemma 3 27BGoogle131K~175$0.08$0.45
16Gemma 3 4BGoogle131K~175$0.05$0.10
17GLM 4.5Zhipu AI131K~175$0.60$2.20
18GLM 4.5 AirZhipu AI131K~175$0.13$0.85
19GLM 4.6VZhipu AI131K~175$0.30$0.90
20gpt-oss-120bOpenAI131K~175$0.15$0.60
21gpt-oss-20bOpenAI131K~175$0.02$0.09
22gpt-oss-20b (batch)OpenAI131K~175$0.02$0.11
23gpt-oss-safeguard-20bOpenAI131K~175$0.07$0.30
24Granite 4.2 8BIBM131K~175$0.06$0.25
25Hunyuan A13B InstructTencent131K~175$0.14$0.57
26Kimi K2 0711Moonshot AI131K~175$0.57$2.30
27Ling 3.0 Flash VLinclusionai131K~175$0.06$0.18
28Llama 3.1 70B InstructMeta131K~175$0.40$0.40
29Llama 3.1 8B InstructMeta131K~175$0.05$0.08
30Llama 3.2 3B InstructMeta131K~175$0.05$0.33
31Llama 3.3 70B InstructMeta131K~175$0.10$0.32
32Ministral 3 3B 2512Mistral AI131K~175$0.10$0.10
33Mistral Large 2407Mistral AI131K~175$2.00$6.00
34Mistral Medium 3Mistral AI131K~175$0.40$2.00
35Mistral Medium 3.1Mistral AI131K~175$0.40$2.00
36Mistral Medium 3.1 (batch)Mistral AI131K~175$0.20$1.00
37Mistral NemoMistral AI131K~175$0.02$0.03
38Muse Glimmer 30Bmeta131K~175$0.30$1.20
39Nano Banana 2 (Gemini 3.1 Flash Image)Google131K~175$0.50$3.00
40Nano Banana Pro (Gemini 3 Pro Image)Google131K~175$2.00$12.00
41Nemotron 3.5 Content SafetyNVIDIA131K~175$0.20$0.20
42Qwen3 14BAlibaba131K~175$0.12$0.24
43Qwen3 235B A22BAlibaba131K~175$0.45$1.82
44Qwen3 235B A22B Thinking 2507Alibaba131K~175$0.23$2.30
45Qwen3 30B A3BAlibaba131K~175$0.12$0.50
46Qwen3 32BAlibaba131K~175$0.08$0.28
47Qwen3 8BAlibaba131K~175$0.12$0.45
48Qwen3 VL 235B A22B ThinkingAlibaba131K~175$0.40$4.00
49Qwen3 VL 32B InstructAlibaba131K~175$0.10$0.42
50Qwen3 VL 8B ThinkingAlibaba131K~175$0.18$2.10
51Solar Pro 3Upstage131K~175$0.15$0.60
52Granite 4.0 MicroIBM131K~175$0.02$0.11
53Command R (08-2024)Cohere128K~171$0.15$0.60
54Command R+ (08-2024)Cohere128K~171$2.50$10.00
55Command R7B (12-2024)Cohere128K~171$0.04$0.15
56GPT AudioOpenAI128K~171$2.50$10.00
57GPT Audio MiniOpenAI128K~171$0.60$2.40
58GPT-4 TurboOpenAI128K~171$10.00$30.00
59GPT-4 Turbo (batch)OpenAI128K~171$5.00$15.00
60GPT-4oOpenAI128K~171$2.50$10.00
61GPT-4o (2024-05-13)OpenAI128K~171$5.00$15.00
62GPT-4o (2024-08-06)OpenAI128K~171$2.50$10.00
63GPT-4o (2024-11-20)OpenAI128K~171$2.50$10.00
64GPT-4o (batch)OpenAI128K~171$1.25$5.00
65GPT-4o-miniOpenAI128K~171$0.15$0.60
66GPT-4o-mini (2024-07-18)OpenAI128K~171$0.15$0.60
67GPT-4o-mini (batch)OpenAI128K~171$0.07$0.30
68GPT-5.2 ChatOpenAI128K~171$1.75$14.00
69Mercury 2Inception128K~171$0.25$0.75
70Mistral LargeMistral AI128K~171$2.00$6.00
71Mistral Small 3.1 24BMistral AI128K~171$0.35$0.55
72Nemotron 3.5 Content Safety (free)NVIDIA128K~171FreeFree
73Nova Micro 1.0Amazon128K~171$0.04$0.14
74Qwen2.5 VL 72B InstructAlibaba128K~171$0.80$1.00
75Schematron V2 Smallinference-net128K~171$0.05$0.23
76Schematron V2 Turboinference-net128K~171$0.03$0.15
77Sonar Deep ResearchPerplexity128K~171$2.00$8.00
78Sonar Reasoning ProPerplexity128K~171$2.00$8.00
79UI-TARS 7B ByteDance128K~171$0.10$0.20

32K–64K Tokens

(27 models)

Moderate context windows suitable for shorter documents, code files, and focused conversations.

#ModelProviderContext WindowPagesInput / 1MOutput / 1M
1SonarPerplexity127K~169$1.00$1.00
2ERNIE 4.5 VL 424B A47B Baidu123K~164$0.42$1.25
3Qwen3 30B A3B Thinking 2507Alibaba82K~109$0.20$2.40
4GLM 4.5VZhipu AI66K~87$0.60$1.80
5LFM2.5-2.6B (free)Liquid AI66K~87FreeFree
6MiniMax M2-herMiniMax66K~87$0.30$1.20
7Mixtral 8x22B InstructMistral AI66K~87$2.00$6.00
8Nano Banana 2 (Gemini 3.1 Flash Image Preview)Google66K~87$0.50$3.00
9Nano Banana 2 Lite (Gemini 3.1 Flash Lite Image)Google66K~87$0.25$1.50
10Nano Banana Pro (Gemini 3 Pro Image Preview)Google66K~87$2.00$12.00
11Reka Flash 3rekaai66K~87$0.10$0.20
12WizardLM-2 8x22BMicrosoft66K~87$0.62$0.62
13R1DeepSeek64K~85$0.70$2.50
14Llama 3.2 1B InstructMeta60K~80$0.03$0.20
15Falcon Arabic 7B InstructTII33K~44FreeFree
16Falcon Mamba 7B InstructTII33K~44FreeFree
17Falcon3 10B InstructTII33K~44FreeFree
18Falcon3 7B InstructTII33K~44FreeFree
19GLM 5.2 (free)Zhipu AI33K~44FreeFree
20Mistral Small 3Mistral AI33K~44$0.05$0.08
21Nano Banana (Gemini 2.5 Flash Image)Google33K~44$0.30$2.50
22Perceptron Mk1perceptron33K~44$0.15$1.50
23Qwen2.5 72B InstructAlibaba33K~44$0.36$0.40
24Qwen2.5 7B InstructAlibaba33K~44$0.10$0.20
25Qwen2.5 Coder 32B InstructAlibaba33K~44$0.66$1.00
26SabaMistral AI33K~44$0.20$0.60
27Voxtral Small 24B 2507Mistral AI33K~44$0.10$0.30

What Is a Context Window and Why Does It Matter?

Context window = working memory

A model's context window is the total number of tokens (roughly words) it can process in a single request. This includes both your input prompt and the model's output. A model with a 128K context window can process about 170 pages of text at once, while a 1M-token model can handle roughly 1,300 pages -- enough for entire books or large codebases.

Larger context enables new use cases

With small context windows (under 32K), you must chunk documents and use retrieval-augmented generation (RAG). Large context models eliminate this complexity for many workloads: analyzing full legal contracts, reviewing entire repositories, summarizing research paper collections, or maintaining very long conversations with full history retained.

Context size vs. effective recall

Not all context is created equal. Some models perform well on "needle in a haystack" tests at their full context length, while others degrade on information retrieval when prompts get very long. The advertised context window is the maximum, but effective performance may vary. Check our leaderboard for quality scores that account for real-world performance.

Cost implications of large context

Using a large context window means sending more tokens per request, which increases cost. For example, filling a 1M-token context at $3/1M input tokens costs $3 per request. For cost-sensitive workloads, consider whether RAG with a smaller context model might be more efficient than filling a large context window end to end.

探索更多

Dive deeper into model capabilities, compare context windows side by side, or see overall rankings across all dimensions.

Frequently Asked Questions

A large context window model can process long inputs - from 128K tokens (about 100 pages) up to 2M tokens (about 1,500 pages). This allows analyzing entire codebases, long documents, or extensive conversation histories in a single request.

Large context models excel when you need the AI to reason across an entire document at once. RAG (Retrieval-Augmented Generation) is better for searching across massive document collections. Large context is simpler to implement but costs more per request.

No. While many models advertise 1M+ token context windows, their effective recall can vary significantly. Some models lose accuracy when important information is buried in the middle of very long contexts. Our rankings account for practical performance, not just advertised limits.

Large Context AI Models - 1M+ Token (2026) | LM Market Cap