Skip to content

Large Context Window AI Models

Compare 411 AI models with context windows of 32K tokens or more. The largest models support 2M tokens -- enough to process entire codebases, books, or hundreds of documents in a single prompt. Data updated hourly.

157
1M+ Tokens
149
200K+ Tokens
78
128K Tokens
27
32K–64K Tokens
2M
Largest Context
411
Models (32K+)
536K
Average Context

1M+ Tokens

(157 models)

The largest context windows available. These models can process entire codebases, full books, or hundreds of documents in a single request.

#ModelProviderContext WindowPagesInput / 1MOutput / 1M
1Grok 4.20xAI2M~3K$1.25$2.50
2Grok 4.20 Multi-AgentxAI2M~3K$1.25$2.50
3DeepSeek V4 Flash 0731DeepSeek1.3M~2K$0.04$0.64
4DeepSeek V4 Flash Latest~deepseek1.3M~2K$0.03$1.00
5GLM 5.3Zhipu AI1.3M~2K$0.65$2.05
6GLM 5.3 FlashZhipu AI1.3M~2K$0.15$0.50
7GLM Flash Latest~z-ai1.3M~2K$0.07$0.25
8GLM Latest~z-ai1.3M~2K$0.65$2.05
9Llama 4 ScoutMeta1.3M~2K$0.10$0.30
10GPT Astra Latest~openai1.1M~1K$10.00$50.00
11GPT Luna Latest~openai1.1M~1K$0.20$1.20
12GPT Sol Latest~openai1.1M~1K$2.00$10.00
13GPT Terra Latest~openai1.1M~1K$2.00$12.00
14GPT-5.4OpenAI1.1M~1K$2.50$15.00
15GPT-5.4 (batch)OpenAI1.1M~1K$1.25$7.50
16GPT-5.4 ProOpenAI1.1M~1K$30.00$180.00
17GPT-5.4 Pro (batch)OpenAI1.1M~1K$15.00$90.00
18GPT-5.5OpenAI1.1M~1K$5.00$30.00
19GPT-5.5 (batch)OpenAI1.1M~1K$2.50$15.00
20GPT-5.5 ProOpenAI1.1M~1K$30.00$180.00
21GPT-5.5 Pro (batch)OpenAI1.1M~1K$15.00$90.00
22GPT-5.6 LunaOpenAI1.1M~1K$0.20$1.20
23GPT-5.6 Luna (batch)OpenAI1.1M~1K$0.10$0.60
24GPT-5.6 Luna ProOpenAI1.1M~1K$0.20$1.20
25GPT-5.6 Luna Pro (batch)OpenAI1.1M~1K$0.10$0.60
26GPT-5.6 SolOpenAI1.1M~1K$2.00$10.00
27GPT-5.6 Sol (batch)OpenAI1.1M~1K$1.00$5.00
28GPT-5.6 Sol ProOpenAI1.1M~1K$2.00$10.00
29GPT-5.6 Sol Pro (batch)OpenAI1.1M~1K$1.00$5.00
30GPT-5.6 TerraOpenAI1.1M~1K$2.00$12.00
31GPT-5.6 Terra (batch)OpenAI1.1M~1K$1.00$6.00
32GPT-5.6 Terra ProOpenAI1.1M~1K$2.00$12.00
33GPT-5.6 Terra Pro (batch)OpenAI1.1M~1K$1.00$6.00
34GPT-6 AstraOpenAI1.1M~1K$10.00$50.00
35GPT-6 Astra (batch)OpenAI1.1M~1K$5.00$25.00
36GPT-6 Astra ProOpenAI1.1M~1K$10.00$50.00
37GPT-6 Astra Pro (batch)OpenAI1.1M~1K$5.00$25.00
38MiMo-V2.5Xiaomi1.1M~1K$0.14$0.28
39MiMo-V2.5-ProXiaomi1.1M~1K$0.43$0.87
40LongCat 2.0Meituan1.0M~1K$0.30$1.20
41DeepSeek Flash Latest~deepseek1.0M~1K$0.12$0.48
42DeepSeek Pro Latest~deepseek1.0M~1K$0.62$1.87
43DeepSeek V4 Flash 0423DeepSeek1.0M~1K$0.05$0.10
44DeepSeek V4 Flash 0731 (batch)DeepSeek1.0M~1K$0.11$0.33
45DeepSeek V4 Flash Vision ExpDeepSeek1.0M~1K$0.22$0.66
46DeepSeek V4 Flash Vision Exp (batch)DeepSeek1.0M~1K$0.11$0.33
47DeepSeek V4 Pro 0423DeepSeek1.0M~1K$0.95$1.91
48DeepSeek V4 Pro 0813DeepSeek1.0M~1K$0.62$1.87
49DeepSeek V4 Pro 0813 (batch)DeepSeek1.0M~1K$0.66$1.98
50DeepSeek V4.1 FlashDeepSeek1.0M~1K$0.15$0.60
51Gemini 2.5 FlashGoogle1.0M~1K$0.30$2.50
52Gemini 2.5 Flash (batch)Google1.0M~1K$0.15$1.25
53Gemini 2.5 Flash LiteGoogle1.0M~1K$0.10$0.40
54Gemini 2.5 Flash Lite (batch)Google1.0M~1K$0.05$0.20
55Gemini 2.5 ProGoogle1.0M~1K$1.25$10.00
56Gemini 2.5 Pro (batch)Google1.0M~1K$0.63$5.00
57Gemini 2.5 Pro Preview 06-05Google1.0M~1K$1.25$10.00
58Gemini 3 Flash PreviewGoogle1.0M~1K$0.50$3.00
59Gemini 3 Flash Preview (batch)Google1.0M~1K$0.25$1.50
60Gemini 3.1 Flash LiteGoogle1.0M~1K$0.25$1.50
61Gemini 3.1 Flash Lite (batch)Google1.0M~1K$0.13$0.75
62Gemini 3.1 Flash Lite PreviewGoogle1.0M~1K$0.25$1.50
63Gemini 3.1 Pro PreviewGoogle1.0M~1K$2.00$12.00
64Gemini 3.1 Pro Preview (batch)Google1.0M~1K$1.00$6.00
65Gemini 3.1 Pro Preview Custom ToolsGoogle1.0M~1K$2.00$12.00
66Gemini 3.5 FlashGoogle1.0M~1K$1.50$9.00
67Gemini 3.5 Flash (batch)Google1.0M~1K$0.75$4.50
68Gemini 3.5 Flash LiteGoogle1.0M~1K$0.30$2.50
69Gemini 3.5 Flash Lite (batch)Google1.0M~1K$0.15$1.25
70Gemini 3.6 FlashGoogle1.0M~1K$0.75$3.75
71Gemini 3.6 Flash (batch)Google1.0M~1K$0.38$1.88
72Gemini 3.7 FlashGoogle1.0M~1K$0.75$3.75
73Gemini 3.7 Flash (batch)Google1.0M~1K$0.38$1.88
74Gemini 3.8 FlashGoogle1.0M~1K$0.75$3.75
75Gemini 3.8 Flash (batch)Google1.0M~1K$0.38$1.88
76Gemini Flash Latest~google1.0M~1K$0.75$3.75
77Gemini Pro Latest~google1.0M~1K$2.00$12.00
78GLM 5.2Zhipu AI1.0M~1K$0.65$2.04
79GLM 5.2 (batch)Zhipu AI1.0M~1K$0.70$2.20
80GLM 5.3 (batch)Zhipu AI1.0M~1K$0.70$2.20
81GLM 5.3 Flash (batch)Zhipu AI1.0M~1K$0.07$0.25
82GLM 5.3 FlashXZhipu AI1.0M~1K$0.37$1.25
83Hy4 previewTencent1.0M~1K$0.83$2.50
84Inklingthinkingmachines1.0M~1K$1.00$4.05
85Inkling (free)thinkingmachines1.0M~1KFreeFree
86Inkling Smallthinkingmachines1.0M~1K$0.45$1.20
87Inkling Small (free)thinkingmachines1.0M~1KFreeFree
88Kimi K3Moonshot AI1.0M~1K$3.00$15.00
89Kimi Latest~moonshotai1.0M~1K$1.50$7.50
90Laguna S 2.1poolside1.0M~1K$0.09$0.18
91Llama 4 MaverickMeta1.0M~1K$0.19$0.65
92Lyria 3 Clip PreviewGoogle1.0M~1KFreeFree
93Lyria 3 Pro PreviewGoogle1.0M~1KFreeFree
94MiMo-V2.6-FlashXiaomi1.0M~1K$0.14$0.28
95MiMo-V2.6-ProXiaomi1.0M~1K$0.43$0.87
96MiMo-V2.6-Pro-UltraSpeedXiaomi1.0M~1K$4.35$8.70
97MiniMax M3MiniMax1.0M~1K$0.30$1.20
98Muse Spark 1.1meta1.0M~1K$1.25$4.25
99Muse Spark 1.2meta1.0M~1K$1.25$4.25
100Muse Spark 1.2 Contributormeta1.0M~1K$0.10$0.20
101Muse Spark 1.3meta1.0M~1K$1.25$4.25
102Muse Spark 1.3 Contributormeta1.0M~1K$0.10$0.20
103Qwen3.8 2.4T A95BAlibaba1.0M~1K$2.00$6.00
104GPT-4.1OpenAI1.0M~1K$2.00$8.00
105GPT-4.1 (batch)OpenAI1.0M~1K$1.00$4.00
106GPT-4.1 MiniOpenAI1.0M~1K$0.40$1.60
107GPT-4.1 Mini (batch)OpenAI1.0M~1K$0.20$0.80
108GPT-4.1 NanoOpenAI1.0M~1K$0.10$0.40
109GPT-4.1 Nano (batch)OpenAI1.0M~1K$0.05$0.20
110Palmyra X5Writer1.0M~1K$0.60$6.00
111MiniMax-01MiniMax1.0M~1K$0.20$1.10
112Claude Fable 5Anthropic1M~1K$10.00$50.00
113Claude Fable 5 (batch)Anthropic1M~1K$5.00$25.00
114Claude Fable 5.1Anthropic1M~1K$10.00$50.00
115Claude Fable 5.1 (batch)Anthropic1M~1K$5.00$25.00
116Claude Fable Latest~anthropic1M~1K$10.00$50.00
117Claude Opus 4.6Anthropic1M~1K$5.00$25.00
118Claude Opus 4.6 (batch)Anthropic1M~1K$2.50$12.50
119Claude Opus 4.7Anthropic1M~1K$5.00$25.00
120Claude Opus 4.7 (batch)Anthropic1M~1K$2.50$12.50
121Claude Opus 4.8Anthropic1M~1K$5.00$25.00
122Claude Opus 4.8 (batch)Anthropic1M~1K$2.50$12.50
123Claude Opus 5Anthropic1M~1K$5.00$25.00
124Claude Opus 5 (batch)Anthropic1M~1K$2.50$12.50
125Claude Opus Latest~anthropic1M~1K$5.00$25.00
126Claude Sonnet 4.5Anthropic1M~1K$3.00$15.00
127Claude Sonnet 4.5 (batch)Anthropic1M~1K$1.50$7.50
128Claude Sonnet 4.6Anthropic1M~1K$3.00$15.00
129Claude Sonnet 4.6 (batch)Anthropic1M~1K$1.50$7.50
130Claude Sonnet 5Anthropic1M~1K$2.00$10.00
131Claude Sonnet 5 (batch)Anthropic1M~1K$1.00$5.00
132Claude Sonnet Latest~anthropic1M~1K$2.00$10.00
133Fugu Maxsakana1M~1K$2.00$6.00
134Fugu Ultrasakana1M~1K$5.00$30.00
135Fugu Ultra v2sakana1M~1K$5.00$30.00
136Grok 4.3xAI1M~1K$1.25$2.50
137Grok 4.3 (batch)xAI1M~1K$1.00$2.00
138MiniMax M1MiniMax1M~1K$0.40$2.20
139Nemotron 3 Ultra (free)NVIDIA1M~1KFreeFree
140Nemotron 3.5 Lightning (free)NVIDIA1M~1KFreeFree
141Nova 2 LiteAmazon1M~1K$0.30$2.50
142Nova Premier 1.0Amazon1M~1K$2.50$12.50
143Qwen Plus 0728Alibaba1M~1K$0.26$0.78
144Qwen-PlusAlibaba1M~1K$0.26$0.78
145Qwen3 Coder FlashAlibaba1M~1K$0.20$0.97
146Qwen3 Coder PlusAlibaba1M~1K$0.65$3.25
147Qwen3.5 Plus 2026-02-15Alibaba1M~1K$0.26$1.56
148Qwen3.5 Plus 2026-04-20Alibaba1M~1K$0.30$1.80
149Qwen3.5-FlashAlibaba1M~1K$0.07$0.26
150Qwen3.6 FlashAlibaba1M~1K$0.19$1.13
151Qwen3.6 PlusAlibaba1M~1K$0.33$1.95
152Qwen3.7 FlashAlibaba1M~1K$0.03$0.13
153Qwen3.7 MaxAlibaba1M~1K$1.48$4.42
154Qwen3.7 PlusAlibaba1M~1K$0.32$1.28
155Qwen3.8 27BAlibaba1M~1K$0.42$3.00
156Qwen3.8 FlashAlibaba1M~1K$0.15$0.47
157Qwen3.8 Max (0902)Alibaba1M~1K$2.00$6.00

200K+ Tokens

(149 models)

Extended context models ideal for long documents, legal contracts, research papers, and multi-file code analysis.

#ModelProviderContext WindowPagesInput / 1MOutput / 1M
1Solar Pro 4Upstage524K~699$0.09$0.36
2Dots3-Note Preview (free)dots-studio512K~683FreeFree
3Grok 4.5xAI500K~667$2.00$6.00
4Grok 4.6xAI500K~667$2.00$6.00
5Grok 4.7xAI500K~667$1.60$4.80
6Grok Latest~x-ai500K~667$1.60$4.80
7GPT Chat LatestOpenAI400K~533$5.00$30.00
8GPT Mini Latest~openai400K~533$0.75$4.50
9GPT-5OpenAI400K~533$1.25$10.00
10GPT-5 (batch)OpenAI400K~533$0.63$5.00
11GPT-5 ImageOpenAI400K~533$10.00$10.00
12GPT-5 Image MiniOpenAI400K~533$2.50$2.00
13GPT-5 MiniOpenAI400K~533$0.25$2.00
14GPT-5 Mini (batch)OpenAI400K~533$0.13$1.00
15GPT-5 NanoOpenAI400K~533$0.05$0.40
16GPT-5 Nano (batch)OpenAI400K~533$0.02$0.20
17GPT-5 ProOpenAI400K~533$15.00$120.00
18GPT-5 Pro (batch)OpenAI400K~533$7.50$60.00
19GPT-5.1OpenAI400K~533$1.25$10.00
20GPT-5.1 (batch)OpenAI400K~533$0.63$5.00
21GPT-5.1-CodexOpenAI400K~533$1.25$10.00
22GPT-5.1-Codex-MaxOpenAI400K~533$1.25$10.00
23GPT-5.1-Codex-MiniOpenAI400K~533$0.25$2.00
24GPT-5.2OpenAI400K~533$1.75$14.00
25GPT-5.2 (batch)OpenAI400K~533$0.88$7.00
26GPT-5.2 ProOpenAI400K~533$21.00$168.00
27GPT-5.2 Pro (batch)OpenAI400K~533$10.50$84.00
28GPT-5.2-CodexOpenAI400K~533$1.75$14.00
29GPT-5.3-CodexOpenAI400K~533$1.75$14.00
30GPT-5.4 MiniOpenAI400K~533$0.75$4.50
31GPT-5.4 Mini (batch)OpenAI400K~533$0.38$2.25
32GPT-5.4 NanoOpenAI400K~533$0.20$1.25
33GPT-5.4 Nano (batch)OpenAI400K~533$0.10$0.63
34Nova Lite 1.0Amazon300K~400$0.06$0.24
35Nova Pro 1.0Amazon300K~400$0.80$3.20
36GPT-5.4 Image 2OpenAI272K~363$8.00$15.00
37Devstral 2 2512Mistral AI262K~350$0.40$2.00
38Falcon-H1-Arabic 34B InstructTII262K~350FreeFree
39Falcon-H1-Arabic 7B InstructTII262K~350FreeFree
40Gemma 4 26B A4B Google262K~350$0.09$0.30
41Gemma 4 26B A4B (free)Google262K~350FreeFree
42Gemma 4 31BGoogle262K~350$0.09$0.34
43Gemma 4 31B (free)Google262K~350FreeFree
44Hy3Tencent262K~350$0.13$0.53
45Hy3 previewTencent262K~350$0.18$0.60
46KAT-Coder-Pro V2.5Kuaishou262K~350$0.74$2.96
47Kimi K2 0905Moonshot AI262K~350$0.60$2.50
48Kimi K2 ThinkingMoonshot AI262K~350$0.60$2.50
49Kimi K2.5Moonshot AI262K~350$0.45$2.25
50Kimi K2.6Moonshot AI262K~350$0.95$4.00
51Kimi K2.7 CodeMoonshot AI262K~350$0.71$3.30
52Laguna S 2.1 (free)poolside262K~350FreeFree
53Laguna XS 2.1poolside262K~350$0.06$0.12
54Laguna XS 2.1 (free)poolside262K~350FreeFree
55Ling 3.0 Flashinclusionai262K~350$0.02$0.06
56Ling 3.0 Flash Fininclusionai262K~350$0.06$0.18
57Ling 3.0 Flash Fin (free)inclusionai262K~350FreeFree
58Ling 3.0 Flash Sante (free)inclusionai262K~350FreeFree
59Ling 3.0 Flash VL (free)inclusionai262K~350FreeFree
60Ministral 3 14B 2512Mistral AI262K~350$0.20$0.20
61Ministral 3 8B 2512Mistral AI262K~350$0.15$0.15
62Ministral 3 8B 2512 (batch)Mistral AI262K~350$0.07$0.07
63Mistral Large 3 2512 (batch)Mistral AI262K~350$0.25$0.75
64Mistral Medium 3.5Mistral AI262K~350$1.50$7.50
65Mistral Medium 3.5 (batch)Mistral AI262K~350$0.75$3.75
66Mistral Small 4Mistral AI262K~350$0.15$0.60
67Mistral Small 4 (batch)Mistral AI262K~350$0.07$0.30
68Nemotron 3 Nano 30B A3BNVIDIA262K~350$0.05$0.20
69Nemotron 3 SuperNVIDIA262K~350$0.08$0.45
70Nemotron 3 Super (free)NVIDIA262K~350FreeFree
71Nemotron 3 UltraNVIDIA262K~350$0.60$2.40
72Nemotron 3.5 LightningNVIDIA262K~350$0.07$0.20
73Paretounbiased262K~350$2.50$7.50
74Qwen3 235B A22B Instruct 2507Alibaba262K~350$0.09$0.35
75Qwen3 30B A3B Instruct 2507Alibaba262K~350$0.05$0.19
76Qwen3 Coder 30B A3B InstructAlibaba262K~350$0.07$0.28
77Qwen3 Coder 480B A35BAlibaba262K~350$0.30$1.00
78Qwen3 Coder NextAlibaba262K~350$0.12$0.80
79Qwen3 MaxAlibaba262K~350$0.78$3.90
80Qwen3 Max ThinkingAlibaba262K~350$0.78$3.90
81Qwen3 Next 80B A3B InstructAlibaba262K~350$0.09$1.10
82Qwen3 Next 80B A3B ThinkingAlibaba262K~350$0.15$1.20
83Qwen3 VL 235B A22B InstructAlibaba262K~350$0.21$1.90
84Qwen3 VL 30B A3B InstructAlibaba262K~350$0.13$0.52
85Qwen3 VL 30B A3B ThinkingAlibaba262K~350$0.20$2.40
86Qwen3 VL 8B InstructAlibaba262K~350$0.12$0.45
87Qwen3.5 397B A17BAlibaba262K~350$0.55$3.50
88Qwen3.5-122B-A10BAlibaba262K~350$0.26$2.08
89Qwen3.5-27BAlibaba262K~350$0.20$1.56
90Qwen3.5-35B-A3BAlibaba262K~350$0.31$1.25
91Qwen3.5-9BAlibaba262K~350$0.10$0.15
92Qwen3.6 27BAlibaba262K~350$0.30$2.00
93Qwen3.6 35B A3BAlibaba262K~350$0.15$1.00
94Qwen3.6 Max PreviewAlibaba262K~350$1.03$6.16
95Qwen3.8 27B (free)Alibaba262K~350FreeFree
96Sakana Namazusakana262K~350$0.95$4.00
97Seed 1.6ByteDance262K~350$0.25$2.00
98Seed 1.6 FlashByteDance262K~350$0.07$0.30
99Seed 2.1 TurboByteDance262K~350$0.50$2.50
100Seed-2.0-CodeByteDance262K~350$0.50$3.00
101Seed-2.0-LiteByteDance262K~350$0.25$2.00
102Seed-2.0-MiniByteDance262K~350$0.10$0.40
103Step 3.5 FlashStepFun262K~350$0.10$0.30
104Step 3.7 FlashStepFun262K~350$0.20$1.15
105Ternary Bonsai 2 27Bprism-ml262K~350$0.07$0.50
106Trinity Large Thinkingarcee-ai262K~350$0.25$0.80
107Mercury 2.5Inception260K~347$0.04$0.15
108Codestral 2508Mistral AI256K~341$0.30$0.90
109Codestral 2508 (batch)Mistral AI256K~341$0.15$0.45
110Command ACohere256K~341$2.50$10.00
111Grok Build 0.1xAI256K~341$1.00$2.00
112Mistral Small 3.2 24BMistral AI256K~341$0.09$0.25
113Nemotron 3 Nano Omni (free)NVIDIA256K~341FreeFree
114North Mini Code (free)Cohere256K~341FreeFree
115GLM 4.6Zhipu AI205K~273$0.43$1.75
116GLM 4.7Zhipu AI205K~273$0.40$1.75
117GLM 5Zhipu AI205K~273$0.60$1.92
118GLM 5.1Zhipu AI205K~273$0.97$3.04
119MiniMax M2MiniMax205K~273$0.26$1.02
120MiniMax M2.1MiniMax205K~273$0.30$1.20
121MiniMax M2.5MiniMax205K~273$0.27$1.08
122MiniMax M2.7MiniMax205K~273$0.30$1.20
123GLM 5 TurboZhipu AI203K~270$1.20$4.00
124GLM 5V TurboZhipu AI203K~270$1.20$4.00
125Claude 3 HaikuAnthropic200K~267$0.25$1.25
126Claude Haiku 4.5Anthropic200K~267$1.00$5.00
127Claude Haiku 4.5 (batch)Anthropic200K~267$0.50$2.50
128Claude Haiku Latest~anthropic200K~267$1.00$5.00
129Claude Opus 4.1Anthropic200K~267$15.00$75.00
130Claude Opus 4.1 (batch)Anthropic200K~267$7.50$37.50
131Claude Opus 4.5Anthropic200K~267$5.00$25.00
132Claude Opus 4.5 (batch)Anthropic200K~267$2.50$12.50
133Claude Sonnet 4Anthropic200K~267$3.00$15.00
134Composer 2Cursor200K~267$0.50$2.50
135Composer 2 FastCursor200K~267$1.50$7.50
136GLM 4.7 FlashZhipu AI200K~267$0.06$0.40
137o1OpenAI200K~267$15.00$60.00
138o1-proOpenAI200K~267$150.00$600.00
139o3OpenAI200K~267$2.00$8.00
140o3 (batch)OpenAI200K~267$1.00$4.00
141o3 MiniOpenAI200K~267$1.10$4.40
142o3 Mini (batch)OpenAI200K~267$0.55$2.20
143o3 Mini HighOpenAI200K~267$1.10$4.40
144o3 ProOpenAI200K~267$20.00$80.00
145o4 MiniOpenAI200K~267$1.10$4.40
146o4 Mini (batch)OpenAI200K~267$0.55$2.20
147o4 Mini HighOpenAI200K~267$1.10$4.40
148Sonar ProPerplexity200K~267$3.00$15.00
149Sonar Pro SearchPerplexity200K~267$3.00$15.00

128K Tokens

(78 models)

The current standard for frontier models. Sufficient for most production use cases, including long conversations and medium-length documents.

#ModelProviderContext WindowPagesInput / 1MOutput / 1M
1DeepSeek V3DeepSeek164K~218$0.32$0.89
2DeepSeek V3 0324DeepSeek164K~218$0.25$1.00
3DeepSeek V3.1DeepSeek164K~218$0.25$0.95
4DeepSeek V3.1 TerminusDeepSeek164K~218$0.27$1.00
5DeepSeek V3.2DeepSeek164K~218$0.27$0.40
6DeepSeek V3.2 ExpDeepSeek164K~218$0.27$0.41
7Llama Guard 4 12BMeta164K~218$0.18$0.18
8R1 0528DeepSeek164K~218$0.50$2.15
9Aion-2.0aion-labs131K~175$0.80$1.60
10Aion-3.0aion-labs131K~175$3.00$6.00
11Aion-3.0-Miniaion-labs131K~175$0.70$1.40
12Falcon-H1-Arabic 3B InstructTII131K~175FreeFree
13Gemma 3 12BGoogle131K~175$0.05$0.15
14Gemma 3 27BGoogle131K~175$0.08$0.45
15Gemma 3 4BGoogle131K~175$0.05$0.10
16GLM 4.5Zhipu AI131K~175$0.60$2.20
17GLM 4.5 AirZhipu AI131K~175$0.13$0.85
18GLM 4.6VZhipu AI131K~175$0.30$0.90
19gpt-oss-120bOpenAI131K~175$0.15$0.60
20gpt-oss-20bOpenAI131K~175$0.03$0.13
21gpt-oss-safeguard-20bOpenAI131K~175$0.07$0.30
22Granite 4.2 8BIBM131K~175$0.06$0.25
23Hunyuan A13B InstructTencent131K~175$0.14$0.57
24Kimi K2 0711Moonshot AI131K~175$0.57$2.30
25Ling 3.0 Flash VLinclusionai131K~175$0.06$0.18
26Llama 3.1 70B InstructMeta131K~175$0.40$0.40
27Llama 3.1 8B InstructMeta131K~175$0.05$0.08
28Llama 3.2 3B InstructMeta131K~175$0.05$0.33
29Llama 3.3 70B InstructMeta131K~175$0.10$0.32
30Ministral 3 3B 2512Mistral AI131K~175$0.10$0.10
31Mistral Large 2407Mistral AI131K~175$2.00$6.00
32Mistral Medium 3Mistral AI131K~175$0.40$2.00
33Mistral Medium 3.1Mistral AI131K~175$0.40$2.00
34Mistral Medium 3.1 (batch)Mistral AI131K~175$0.20$1.00
35Mistral NemoMistral AI131K~175$0.02$0.03
36Muse Glimmer 30Bmeta131K~175$0.30$1.20
37Muse Glimmer 30B (batch)meta131K~175$0.17$0.75
38Nano Banana 2 (Gemini 3.1 Flash Image)Google131K~175$0.50$3.00
39Nano Banana Pro (Gemini 3 Pro Image)Google131K~175$2.00$12.00
40Nemotron 3.5 Content SafetyNVIDIA131K~175$0.20$0.20
41Qwen3 14BAlibaba131K~175$0.12$0.24
42Qwen3 235B A22BAlibaba131K~175$0.45$1.82
43Qwen3 235B A22B Thinking 2507Alibaba131K~175$0.23$2.30
44Qwen3 30B A3BAlibaba131K~175$0.12$0.50
45Qwen3 32BAlibaba131K~175$0.08$0.28
46Qwen3 8BAlibaba131K~175$0.12$0.45
47Qwen3 VL 235B A22B ThinkingAlibaba131K~175$0.40$4.00
48Qwen3 VL 32B InstructAlibaba131K~175$0.10$0.42
49Qwen3 VL 8B ThinkingAlibaba131K~175$0.18$2.10
50Solar Pro 3Upstage131K~175$0.15$0.60
51Granite 4.0 MicroIBM131K~175$0.02$0.11
52Command R (08-2024)Cohere128K~171$0.15$0.60
53Command R+ (08-2024)Cohere128K~171$2.50$10.00
54Command R7B (12-2024)Cohere128K~171$0.04$0.15
55GPT AudioOpenAI128K~171$2.50$10.00
56GPT Audio MiniOpenAI128K~171$0.60$2.40
57GPT-4 TurboOpenAI128K~171$10.00$30.00
58GPT-4 Turbo (batch)OpenAI128K~171$5.00$15.00
59GPT-4oOpenAI128K~171$2.50$10.00
60GPT-4o (2024-05-13)OpenAI128K~171$5.00$15.00
61GPT-4o (2024-08-06)OpenAI128K~171$2.50$10.00
62GPT-4o (2024-11-20)OpenAI128K~171$2.50$10.00
63GPT-4o (batch)OpenAI128K~171$1.25$5.00
64GPT-4o-miniOpenAI128K~171$0.15$0.60
65GPT-4o-mini (2024-07-18)OpenAI128K~171$0.15$0.60
66GPT-4o-mini (batch)OpenAI128K~171$0.07$0.30
67GPT-5.2 ChatOpenAI128K~171$1.75$14.00
68Mercury 2Inception128K~171$0.25$0.75
69Mistral LargeMistral AI128K~171$2.00$6.00
70Mistral Small 3.1 24BMistral AI128K~171$0.35$0.55
71Nemotron 3.5 Content Safety (free)NVIDIA128K~171FreeFree
72Nova Micro 1.0Amazon128K~171$0.04$0.14
73Qwen2.5 VL 72B InstructAlibaba128K~171$0.80$1.00
74Schematron V2 Smallinference-net128K~171$0.05$0.23
75Schematron V2 Turboinference-net128K~171$0.03$0.15
76Sonar Deep ResearchPerplexity128K~171$2.00$8.00
77Sonar Reasoning ProPerplexity128K~171$2.00$8.00
78UI-TARS 7B ByteDance128K~171$0.10$0.20

32K–64K Tokens

(27 models)

Moderate context windows suitable for shorter documents, code files, and focused conversations.

#ModelProviderContext WindowPagesInput / 1MOutput / 1M
1SonarPerplexity127K~169$1.00$1.00
2ERNIE 4.5 VL 424B A47B Baidu123K~164$0.42$1.25
3Qwen3 30B A3B Thinking 2507Alibaba82K~109$0.20$2.40
4GLM 4.5VZhipu AI66K~87$0.60$1.80
5LFM2.5-2.6B (free)Liquid AI66K~87FreeFree
6MiniMax M2-herMiniMax66K~87$0.30$1.20
7Mixtral 8x22B InstructMistral AI66K~87$2.00$6.00
8Nano Banana 2 (Gemini 3.1 Flash Image Preview)Google66K~87$0.50$3.00
9Nano Banana 2 Lite (Gemini 3.1 Flash Lite Image)Google66K~87$0.25$1.50
10Nano Banana Pro (Gemini 3 Pro Image Preview)Google66K~87$2.00$12.00
11Reka Flash 3rekaai66K~87$0.10$0.20
12WizardLM-2 8x22BMicrosoft66K~87$0.62$0.62
13R1DeepSeek64K~85$0.70$2.50
14Llama 3.2 1B InstructMeta60K~80$0.03$0.20
15Falcon Arabic 7B InstructTII33K~44FreeFree
16Falcon Mamba 7B InstructTII33K~44FreeFree
17Falcon3 10B InstructTII33K~44FreeFree
18Falcon3 7B InstructTII33K~44FreeFree
19GLM 5.2 (free)Zhipu AI33K~44FreeFree
20Mistral Small 3Mistral AI33K~44$0.05$0.08
21Nano Banana (Gemini 2.5 Flash Image)Google33K~44$0.30$2.50
22Perceptron Mk1perceptron33K~44$0.15$1.50
23Qwen2.5 72B InstructAlibaba33K~44$0.36$0.40
24Qwen2.5 7B InstructAlibaba33K~44$0.10$0.20
25Qwen2.5 Coder 32B InstructAlibaba33K~44$0.66$1.00
26SabaMistral AI33K~44$0.20$0.60
27Voxtral Small 24B 2507Mistral AI33K~44$0.10$0.30

What Is a Context Window and Why Does It Matter?

Context window = working memory

A model's context window is the total number of tokens (roughly words) it can process in a single request. This includes both your input prompt and the model's output. A model with a 128K context window can process about 170 pages of text at once, while a 1M-token model can handle roughly 1,300 pages -- enough for entire books or large codebases.

Larger context enables new use cases

With small context windows (under 32K), you must chunk documents and use retrieval-augmented generation (RAG). Large context models eliminate this complexity for many workloads: analyzing full legal contracts, reviewing entire repositories, summarizing research paper collections, or maintaining very long conversations with full history retained.

Context size vs. effective recall

Not all context is created equal. Some models perform well on "needle in a haystack" tests at their full context length, while others degrade on information retrieval when prompts get very long. The advertised context window is the maximum, but effective performance may vary. Check our leaderboard for quality scores that account for real-world performance.

Cost implications of large context

Using a large context window means sending more tokens per request, which increases cost. For example, filling a 1M-token context at $3/1M input tokens costs $3 per request. For cost-sensitive workloads, consider whether RAG with a smaller context model might be more efficient than filling a large context window end to end.

Explore More

Dive deeper into model capabilities, compare context windows side by side, or see overall rankings across all dimensions.

Frequently Asked Questions

A large context window model can process long inputs - from 128K tokens (about 100 pages) up to 2M tokens (about 1,500 pages). This allows analyzing entire codebases, long documents, or extensive conversation histories in a single request.

Large context models excel when you need the AI to reason across an entire document at once. RAG (Retrieval-Augmented Generation) is better for searching across massive document collections. Large context is simpler to implement but costs more per request.

No. While many models advertise 1M+ token context windows, their effective recall can vary significantly. Some models lose accuracy when important information is buried in the middle of very long contexts. Our rankings account for practical performance, not just advertised limits.

Large Context AI Models - 1M+ Token (2026) | LM Market Cap