Skip to content

Large Context Window AI Models

Compare 370 AI models with context windows of 32K tokens or more. The largest models support 2M tokens -- enough to process entire codebases, books, or hundreds of documents in a single prompt. Data updated hourly.

119
1M+ Tokens
149
200K+ Tokens
76
128K Tokens
26
32K–64K Tokens
2M
Largest Context
370
Models (32K+)
483K
Average Context

1M+ Tokens

(119 models)

The largest context windows available. These models can process entire codebases, full books, or hundreds of documents in a single request.

#ModelProviderContext WindowPagesInput / 1MOutput / 1M
1Grok 4.20xAI2M~3K$1.25$2.50
2Grok 4.20 Multi-AgentxAI2M~3K$1.25$2.50
3Llama 4 ScoutMeta1.3M~2K$0.10$0.30
4GPT-5.4OpenAI1.1M~1K$2.50$15.00
5GPT-5.4 (batch)OpenAI1.1M~1K$1.25$7.50
6GPT-5.4 ProOpenAI1.1M~1K$30.00$180.00
7GPT-5.4 Pro (batch)OpenAI1.1M~1K$15.00$90.00
8GPT-5.5OpenAI1.1M~1K$5.00$30.00
9GPT-5.5 (batch)OpenAI1.1M~1K$2.50$15.00
10GPT-5.5 ProOpenAI1.1M~1K$30.00$180.00
11GPT-5.5 Pro (batch)OpenAI1.1M~1K$15.00$90.00
12GPT-5.6 LunaOpenAI1.1M~1K$0.10$0.60
13GPT-5.6 Luna (batch)OpenAI1.1M~1K$0.10$0.60
14GPT-5.6 Luna ProOpenAI1.1M~1K$0.10$0.60
15GPT-5.6 Luna Pro (batch)OpenAI1.1M~1K$0.10$0.60
16GPT-5.6 SolOpenAI1.1M~1K$5.00$30.00
17GPT-5.6 Sol (batch)OpenAI1.1M~1K$2.50$15.00
18GPT-5.6 Sol ProOpenAI1.1M~1K$5.00$30.00
19GPT-5.6 Sol Pro (batch)OpenAI1.1M~1K$2.50$15.00
20GPT-5.6 TerraOpenAI1.1M~1K$1.00$6.00
21GPT-5.6 Terra (batch)OpenAI1.1M~1K$1.00$6.00
22GPT-5.6 Terra ProOpenAI1.1M~1K$1.00$6.00
23GPT-5.6 Terra Pro (batch)OpenAI1.1M~1K$1.00$6.00
24MiMo-V2.5Xiaomi1.1M~1K$0.14$0.28
25MiMo-V2.5-ProXiaomi1.1M~1K$0.43$0.87
26OpenAI GPT Latest~openai1.1M~1K$5.00$30.00
27LongCat 2.0Meituan1.0M~1K$0.30$1.20
28DeepSeek V4 Flash 0423DeepSeek1.0M~1K$0.14$0.28
29DeepSeek V4 Flash 0731DeepSeek1.0M~1K$0.09$0.18
30DeepSeek V4 Flash Latest~deepseek1.0M~1K$0.09$0.18
31DeepSeek V4 ProDeepSeek1.0M~1K$0.43$0.87
32Gemini 2.5 FlashGoogle1.0M~1K$0.30$2.50
33Gemini 2.5 Flash (batch)Google1.0M~1K$0.15$1.25
34Gemini 2.5 Flash LiteGoogle1.0M~1K$0.10$0.40
35Gemini 2.5 Flash Lite (batch)Google1.0M~1K$0.05$0.20
36Gemini 2.5 ProGoogle1.0M~1K$1.25$10.00
37Gemini 2.5 Pro (batch)Google1.0M~1K$0.63$5.00
38Gemini 2.5 Pro Preview 05-06Google1.0M~1K$1.25$10.00
39Gemini 2.5 Pro Preview 06-05Google1.0M~1K$1.25$10.00
40Gemini 3 Flash PreviewGoogle1.0M~1K$0.50$3.00
41Gemini 3 Flash Preview (batch)Google1.0M~1K$0.25$1.50
42Gemini 3.1 Flash LiteGoogle1.0M~1K$0.25$1.50
43Gemini 3.1 Flash Lite (batch)Google1.0M~1K$0.13$0.75
44Gemini 3.1 Flash Lite PreviewGoogle1.0M~1K$0.25$1.50
45Gemini 3.1 Pro PreviewGoogle1.0M~1K$2.00$12.00
46Gemini 3.1 Pro Preview (batch)Google1.0M~1K$1.00$6.00
47Gemini 3.1 Pro Preview Custom ToolsGoogle1.0M~1K$2.00$12.00
48Gemini 3.5 FlashGoogle1.0M~1K$1.50$9.00
49Gemini 3.5 Flash (batch)Google1.0M~1K$0.75$4.50
50Gemini 3.5 Flash LiteGoogle1.0M~1K$0.30$2.50
51Gemini 3.5 Flash Lite (batch)Google1.0M~1K$0.15$1.25
52Gemini 3.6 FlashGoogle1.0M~1K$1.50$7.50
53Gemini 3.6 Flash (batch)Google1.0M~1K$0.75$3.75
54GLM 5.2Zhipu AI1.0M~1K$0.24$0.74
55Google Gemini Flash Latest~google1.0M~1K$1.50$7.50
56Google Gemini Pro Latest~google1.0M~1K$2.00$12.00
57Inklingthinkingmachines1.0M~1K$0.95$4.05
58Kimi K3Moonshot AI1.0M~1K$3.00$15.00
59Laguna S 2.1poolside1.0M~1K$0.09$0.18
60Llama 4 MaverickMeta1.0M~1K$0.20$0.80
61Llama Guard 4 12BMeta1.0M~1K$0.18$0.18
62Lyria 3 Clip PreviewGoogle1.0M~1KFreeFree
63Lyria 3 Pro PreviewGoogle1.0M~1KFreeFree
64MiniMax M3MiniMax1.0M~1K$0.30$1.20
65MoonshotAI Kimi Latest~moonshotai1.0M~1K$2.50$14.00
66Muse Spark 1.1meta1.0M~1K$1.25$4.25
67Muse Spark 1.2meta1.0M~1K$1.25$4.25
68GPT-4.1OpenAI1.0M~1K$2.00$8.00
69GPT-4.1 (batch)OpenAI1.0M~1K$1.00$4.00
70GPT-4.1 MiniOpenAI1.0M~1K$0.40$1.60
71GPT-4.1 Mini (batch)OpenAI1.0M~1K$0.20$0.80
72GPT-4.1 NanoOpenAI1.0M~1K$0.10$0.40
73GPT-4.1 Nano (batch)OpenAI1.0M~1K$0.05$0.20
74Palmyra X5Writer1.0M~1K$0.60$6.00
75MiniMax-01MiniMax1.0M~1K$0.20$1.10
76Anthropic Claude Sonnet Latest~anthropic1M~1K$2.00$10.00
77Claude Fable 5Anthropic1M~1K$10.00$50.00
78Claude Fable 5 (batch)Anthropic1M~1K$5.00$25.00
79Claude Fable Latest~anthropic1M~1K$10.00$50.00
80Claude Opus 4.6Anthropic1M~1K$5.00$25.00
81Claude Opus 4.6 (batch)Anthropic1M~1K$2.50$12.50
82Claude Opus 4.7Anthropic1M~1K$5.00$25.00
83Claude Opus 4.7 (batch)Anthropic1M~1K$2.50$12.50
84Claude Opus 4.7 (Fast)Anthropic1M~1K$30.00$150.00
85Claude Opus 4.8Anthropic1M~1K$5.00$25.00
86Claude Opus 4.8 (batch)Anthropic1M~1K$2.50$12.50
87Claude Opus 4.8 (Fast)Anthropic1M~1K$10.00$50.00
88Claude Opus 5Anthropic1M~1K$5.00$25.00
89Claude Opus 5 (batch)Anthropic1M~1K$2.50$12.50
90Claude Opus 5 (Fast)Anthropic1M~1K$10.00$50.00
91Claude Opus Latest~anthropic1M~1K$5.00$25.00
92Claude Sonnet 4Anthropic1M~1K$3.00$15.00
93Claude Sonnet 4.5Anthropic1M~1K$3.00$15.00
94Claude Sonnet 4.5 (batch)Anthropic1M~1K$1.50$7.50
95Claude Sonnet 4.6Anthropic1M~1K$3.00$15.00
96Claude Sonnet 4.6 (batch)Anthropic1M~1K$1.50$7.50
97Claude Sonnet 5Anthropic1M~1K$2.00$10.00
98Claude Sonnet 5 (batch)Anthropic1M~1K$1.00$5.00
99Fugu Ultrasakana1M~1K$5.00$30.00
100Grok 4.3xAI1M~1K$1.25$2.50
101MiniMax M1MiniMax1M~1K$0.55$2.20
102Nemotron 3 SuperNVIDIA1M~1K$0.08$0.40
103Nemotron 3 Ultra (free)NVIDIA1M~1KFreeFree
104Nova 2 LiteAmazon1M~1K$0.30$2.50
105Nova Premier 1.0Amazon1M~1K$2.50$12.50
106Qwen Plus 0728Alibaba1M~1K$0.26$0.78
107Qwen Plus 0728 (thinking)Alibaba1M~1K$0.40$1.20
108Qwen-PlusAlibaba1M~1K$0.26$0.78
109Qwen3 Coder FlashAlibaba1M~1K$0.20$0.97
110Qwen3 Coder PlusAlibaba1M~1K$0.65$3.25
111Qwen3.5 Plus 2026-02-15Alibaba1M~1K$0.26$1.56
112Qwen3.5 Plus 2026-04-20Alibaba1M~1K$0.30$1.80
113Qwen3.5-FlashAlibaba1M~1K$0.07$0.26
114Qwen3.6 FlashAlibaba1M~1K$0.19$1.13
115Qwen3.6 PlusAlibaba1M~1K$0.33$1.95
116Qwen3.7 FlashAlibaba1M~1K$0.03$0.13
117Qwen3.7 MaxAlibaba1M~1K$1.48$4.42
118Qwen3.7 PlusAlibaba1M~1K$0.32$1.28
119Qwen3.8 MaxAlibaba1M~1K$2.00$6.00

200K+ Tokens

(149 models)

Extended context models ideal for long documents, legal contracts, research papers, and multi-file code analysis.

#ModelProviderContext WindowPagesInput / 1MOutput / 1M
1Inkling (batch)thinkingmachines524K~699$0.50$2.02
2Inkling Smallthinkingmachines524K~699$0.45$1.20
3MiniMax M3 (batch)MiniMax524K~699$0.15$0.60
4Nemotron 3 UltraNVIDIA512K~683$0.60$3.60
5Nemotron 3 Ultra (batch)NVIDIA512K~683$0.30$1.80
6GLM 5.2 (batch)Zhipu AI512K~683$0.70$2.20
7Grok 4.5xAI500K~667$2.00$6.00
8Grok Latest~x-ai500K~667$2.00$6.00
9GPT Chat LatestOpenAI400K~533$5.00$30.00
10GPT-5OpenAI400K~533$1.25$10.00
11GPT-5 (batch)OpenAI400K~533$0.63$5.00
12GPT-5 Codex (batch)OpenAI400K~533$0.63$5.00
13GPT-5 ImageOpenAI400K~533$10.00$10.00
14GPT-5 Image MiniOpenAI400K~533$2.50$2.00
15GPT-5 MiniOpenAI400K~533$0.25$2.00
16GPT-5 Mini (batch)OpenAI400K~533$0.13$1.00
17GPT-5 NanoOpenAI400K~533$0.05$0.40
18GPT-5 Nano (batch)OpenAI400K~533$0.02$0.20
19GPT-5 ProOpenAI400K~533$15.00$120.00
20GPT-5 Pro (batch)OpenAI400K~533$7.50$60.00
21GPT-5.1OpenAI400K~533$1.25$10.00
22GPT-5.1 (batch)OpenAI400K~533$0.63$5.00
23GPT-5.1-CodexOpenAI400K~533$1.25$10.00
24GPT-5.1-Codex-MaxOpenAI400K~533$1.25$10.00
25GPT-5.1-Codex-MiniOpenAI400K~533$0.25$2.00
26GPT-5.2OpenAI400K~533$1.75$14.00
27GPT-5.2 (batch)OpenAI400K~533$0.88$7.00
28GPT-5.2 ProOpenAI400K~533$21.00$168.00
29GPT-5.2 Pro (batch)OpenAI400K~533$10.50$84.00
30GPT-5.2-CodexOpenAI400K~533$1.75$14.00
31GPT-5.3-CodexOpenAI400K~533$1.75$14.00
32GPT-5.4 MiniOpenAI400K~533$0.75$4.50
33GPT-5.4 Mini (batch)OpenAI400K~533$0.38$2.25
34GPT-5.4 NanoOpenAI400K~533$0.20$1.25
35GPT-5.4 Nano (batch)OpenAI400K~533$0.10$0.63
36OpenAI GPT Mini Latest~openai400K~533$0.75$4.50
37Nova Lite 1.0Amazon300K~400$0.06$0.24
38Nova Pro 1.0Amazon300K~400$0.80$3.20
39GPT-5.4 Image 2OpenAI272K~363$8.00$15.00
40Falcon-H1-Arabic 34B InstructTII262K~350FreeFree
41Falcon-H1-Arabic 7B InstructTII262K~350FreeFree
42Gemma 3 27BGoogle262K~350$0.08$0.45
43Gemma 4 26B A4B Google262K~350$0.07$0.34
44Gemma 4 26B A4B (free)Google262K~350FreeFree
45Gemma 4 31BGoogle262K~350$0.10$0.34
46Gemma 4 31B (free)Google262K~350FreeFree
47Hy3Tencent262K~350$0.13$0.53
48Hy3 previewTencent262K~350$0.06$0.21
49KAT-Coder-Pro V2Kuaishou262K~350$0.30$1.20
50Kimi K2 0905Moonshot AI262K~350$0.60$2.50
51Kimi K2 ThinkingMoonshot AI262K~350$0.60$2.50
52Kimi K2.5Moonshot AI262K~350$0.57$2.85
53Kimi K2.6Moonshot AI262K~350$0.58$2.44
54Kimi K2.7 CodeMoonshot AI262K~350$0.70$3.50
55Kimi K2.7 Code (batch)Moonshot AI262K~350$0.47$2.00
56Laguna S 2.1 (free)poolside262K~350FreeFree
57Laguna XS 2.1poolside262K~350$0.06$0.12
58Laguna XS 2.1 (free)poolside262K~350FreeFree
59Ling 3.0 Tiny (free)inclusionai262K~350FreeFree
60Ling-2.6-1Tinclusionai262K~350$0.07$0.63
61Ling-2.6-flashinclusionai262K~350$0.01$0.03
62Ling-3.0-flashinclusionai262K~350$0.02$0.06
63Ministral 3 14B 2512Mistral AI262K~350$0.20$0.20
64Ministral 3 8B 2512Mistral AI262K~350$0.15$0.15
65Mistral Large 3 2512Mistral AI262K~350$0.50$1.50
66Mistral Medium 3.5Mistral AI262K~350$1.50$7.50
67Mistral Small 4Mistral AI262K~350$0.15$0.60
68Nemotron 3 Nano 30B A3BNVIDIA262K~350$0.05$0.20
69Nemotron 3 Super (free)NVIDIA262K~350FreeFree
70Qwen3 235B A22B Instruct 2507Alibaba262K~350$0.09$0.55
71Qwen3 235B A22B Thinking 2507Alibaba262K~350$0.23$2.30
72Qwen3 30B A3B Instruct 2507Alibaba262K~350$0.05$0.19
73Qwen3 Coder 30B A3B InstructAlibaba262K~350$0.07$0.27
74Qwen3 Coder 480B A35BAlibaba262K~350$0.30$1.00
75Qwen3 Coder NextAlibaba262K~350$0.12$0.80
76Qwen3 MaxAlibaba262K~350$0.78$3.90
77Qwen3 Max ThinkingAlibaba262K~350$0.78$3.90
78Qwen3 Next 80B A3B InstructAlibaba262K~350$0.09$1.10
79Qwen3 Next 80B A3B ThinkingAlibaba262K~350$0.15$1.20
80Qwen3 VL 235B A22B InstructAlibaba262K~350$0.21$1.90
81Qwen3 VL 30B A3B InstructAlibaba262K~350$0.15$0.60
82Qwen3 VL 30B A3B ThinkingAlibaba262K~350$0.20$2.40
83Qwen3 VL 8B InstructAlibaba262K~350$0.12$0.45
84Qwen3.5 397B A17BAlibaba262K~350$0.39$2.34
85Qwen3.5-122B-A10BAlibaba262K~350$0.29$2.40
86Qwen3.5-27BAlibaba262K~350$0.20$1.56
87Qwen3.5-35B-A3BAlibaba262K~350$0.14$1.00
88Qwen3.5-9BAlibaba262K~350$0.10$0.15
89Qwen3.6 27BAlibaba262K~350$0.60$3.60
90Qwen3.6 35B A3BAlibaba262K~350$0.15$1.00
91Qwen3.6 Max PreviewAlibaba262K~350$1.03$6.16
92Ring-2.6-1Tinclusionai262K~350$0.07$0.63
93Seed 1.6ByteDance262K~350$0.25$2.00
94Seed 1.6 FlashByteDance262K~350$0.07$0.30
95Seed-2.0-LiteByteDance262K~350$0.25$2.00
96Seed-2.0-MiniByteDance262K~350$0.10$0.40
97Step 3.5 FlashStepFun262K~350$0.10$0.30
98Step 3.7 FlashStepFun262K~350$0.20$1.15
99Trinity Large Thinkingarcee-ai262K~350$0.22$0.85
100Codestral 2508Mistral AI256K~341$0.30$0.90
101Command ACohere256K~341$2.50$10.00
102Grok Build 0.1xAI256K~341$1.00$2.00
103Jamba Large 1.7AI21 Labs256K~341$2.00$8.00
104KAT-Coder-Air V2.5Kuaishou256K~341$0.15$0.60
105KAT-Coder-Pro V2.5Kuaishou256K~341$0.74$2.96
106Mistral Small 3.2 24BMistral AI256K~341$0.09$0.25
107Nemotron 3 Nano 30B A3B (free)NVIDIA256K~341FreeFree
108Nemotron 3 Nano Omni (free)NVIDIA256K~341FreeFree
109North Mini Code (free)Cohere256K~341FreeFree
110GLM 4.6Zhipu AI205K~273$0.50$2.00
111GLM 4.7Zhipu AI205K~273$0.40$1.75
112GLM 5Zhipu AI205K~273$0.95$2.55
113GLM 5.1Zhipu AI205K~273$0.95$2.99
114MiniMax M2MiniMax205K~273$0.26$1.02
115MiniMax M2.1MiniMax205K~273$0.30$1.20
116MiniMax M2.5MiniMax205K~273$0.22$0.90
117MiniMax M2.7MiniMax205K~273$0.27$1.08
118GLM 4.7 FlashZhipu AI203K~270$0.06$0.40
119GLM 5 TurboZhipu AI203K~270$1.20$4.00
120GLM 5V TurboZhipu AI203K~270$1.20$4.00
121Anthropic Claude Haiku Latest~anthropic200K~267$1.00$5.00
122Claude 3 HaikuAnthropic200K~267$0.25$1.25
123Claude Haiku 4.5Anthropic200K~267$1.00$5.00
124Claude Haiku 4.5 (batch)Anthropic200K~267$0.50$2.50
125Claude Opus 4Anthropic200K~267$15.00$75.00
126Claude Opus 4.1Anthropic200K~267$15.00$75.00
127Claude Opus 4.1 (batch)Anthropic200K~267$7.50$37.50
128Claude Opus 4.5Anthropic200K~267$5.00$25.00
129Claude Opus 4.5 (batch)Anthropic200K~267$2.50$12.50
130Composer 2Cursor200K~267$0.50$2.50
131Composer 2 FastCursor200K~267$1.50$7.50
132o1OpenAI200K~267$15.00$60.00
133o1 (batch)OpenAI200K~267$7.50$30.00
134o1-proOpenAI200K~267$150.00$600.00
135o1-pro (batch)OpenAI200K~267$75.00$300.00
136o3OpenAI200K~267$2.00$8.00
137o3 (batch)OpenAI200K~267$1.00$4.00
138o3 MiniOpenAI200K~267$1.10$4.40
139o3 Mini (batch)OpenAI200K~267$0.55$2.20
140o3 Mini HighOpenAI200K~267$1.10$4.40
141o3 Mini High (batch)OpenAI200K~267$0.55$2.20
142o3 ProOpenAI200K~267$20.00$80.00
143o3 Pro (batch)OpenAI200K~267$10.00$40.00
144o4 MiniOpenAI200K~267$1.10$4.40
145o4 Mini (batch)OpenAI200K~267$0.55$2.20
146o4 Mini HighOpenAI200K~267$1.10$4.40
147o4 Mini High (batch)OpenAI200K~267$0.55$2.20
148Sonar ProPerplexity200K~267$3.00$15.00
149Sonar Pro SearchPerplexity200K~267$3.00$15.00

128K Tokens

(76 models)

The current standard for frontier models. Sufficient for most production use cases, including long conversations and medium-length documents.

#ModelProviderContext WindowPagesInput / 1MOutput / 1M
1DeepSeek V3DeepSeek164K~218$0.26$1.03
2DeepSeek V3 0324DeepSeek164K~218$0.27$1.12
3DeepSeek V3.1DeepSeek164K~218$0.25$0.95
4DeepSeek V3.1 TerminusDeepSeek164K~218$0.27$1.00
5DeepSeek V3.2DeepSeek164K~218$0.26$0.38
6DeepSeek V3.2 ExpDeepSeek164K~218$0.27$0.41
7R1DeepSeek164K~218$0.70$2.50
8R1 0528DeepSeek164K~218$0.50$2.15
9Aion-2.0aion-labs131K~175$0.80$1.60
10Aion-3.0aion-labs131K~175$3.00$6.00
11Aion-3.0-Miniaion-labs131K~175$0.70$1.40
12Falcon-H1-Arabic 3B InstructTII131K~175FreeFree
13Gemma 3 12BGoogle131K~175$0.05$0.15
14Gemma 3 4BGoogle131K~175$0.05$0.10
15GLM 4.5Zhipu AI131K~175$0.60$2.20
16GLM 4.5 AirZhipu AI131K~175$0.13$0.85
17GLM 4.6VZhipu AI131K~175$0.30$0.90
18gpt-oss-120bOpenAI131K~175$0.04$0.17
19gpt-oss-20bOpenAI131K~175$0.03$0.13
20gpt-oss-20b (free)OpenAI131K~175FreeFree
21gpt-oss-safeguard-20bOpenAI131K~175$0.07$0.30
22Granite 4.1 8BIBM131K~175$0.05$0.10
23Hunyuan A13B InstructTencent131K~175$0.14$0.57
24Kimi K2 0711Moonshot AI131K~175$0.57$2.30
25Llama 3.1 70B InstructMeta131K~175$0.40$0.40
26Llama 3.1 8B InstructMeta131K~175$0.05$0.08
27Llama 3.2 3B InstructMeta131K~175$0.05$0.33
28Llama 3.3 70B InstructMeta131K~175$0.10$0.32
29Ministral 3 3B 2512Mistral AI131K~175$0.10$0.10
30Mistral Large 2407Mistral AI131K~175$2.00$6.00
31Mistral Medium 3Mistral AI131K~175$0.40$2.00
32Mistral Medium 3.1Mistral AI131K~175$0.40$2.00
33Mistral NemoMistral AI131K~175$0.02$0.03
34Nano Banana 2 (Gemini 3.1 Flash Image)Google131K~175$0.50$3.00
35Nano Banana Pro (Gemini 3 Pro Image)Google131K~175$2.00$12.00
36Qwen3 14BAlibaba131K~175$0.23$0.91
37Qwen3 235B A22BAlibaba131K~175$0.45$1.82
38Qwen3 30B A3BAlibaba131K~175$0.12$0.50
39Qwen3 32BAlibaba131K~175$0.08$0.28
40Qwen3 8BAlibaba131K~175$0.12$0.45
41Qwen3 VL 235B A22B ThinkingAlibaba131K~175$0.40$4.00
42Qwen3 VL 32B InstructAlibaba131K~175$0.10$0.42
43Qwen3 VL 8B ThinkingAlibaba131K~175$0.18$2.10
44Solar Pro 3Upstage131K~175$0.15$0.60
45Virtuoso Largearcee-ai131K~175$0.75$1.20
46Granite 4.0 MicroIBM131K~175$0.02$0.11
47Cogito v2.1 671Bdeepcogito128K~171$1.25$1.25
48Command R (08-2024)Cohere128K~171$0.15$0.60
49Command R+ (08-2024)Cohere128K~171$2.50$10.00
50Command R7B (12-2024)Cohere128K~171$0.04$0.15
51GPT AudioOpenAI128K~171$2.50$10.00
52GPT Audio MiniOpenAI128K~171$0.60$2.40
53GPT-4 TurboOpenAI128K~171$10.00$30.00
54GPT-4 Turbo (batch)OpenAI128K~171$5.00$15.00
55GPT-4 Turbo PreviewOpenAI128K~171$10.00$30.00
56GPT-4oOpenAI128K~171$2.50$10.00
57GPT-4o (2024-05-13)OpenAI128K~171$5.00$15.00
58GPT-4o (2024-08-06)OpenAI128K~171$2.50$10.00
59GPT-4o (2024-11-20)OpenAI128K~171$2.50$10.00
60GPT-4o (batch)OpenAI128K~171$1.25$5.00
61GPT-4o-miniOpenAI128K~171$0.15$0.60
62GPT-4o-mini (2024-07-18)OpenAI128K~171$0.15$0.60
63GPT-4o-mini (batch)OpenAI128K~171$0.07$0.30
64GPT-5.2 ChatOpenAI128K~171$1.75$14.00
65GPT-5.3 ChatOpenAI128K~171$1.75$14.00
66Mercury 2Inception128K~171$0.25$0.75
67Mistral LargeMistral AI128K~171$2.00$6.00
68Mistral Small 3.1 24BMistral AI128K~171$0.35$0.55
69Nemotron 3.5 Content Safety (free)NVIDIA128K~171FreeFree
70Nemotron Nano 12B 2 VL (free)NVIDIA128K~171FreeFree
71Nemotron Nano 9B V2 (free)NVIDIA128K~171FreeFree
72Nova Micro 1.0Amazon128K~171$0.04$0.14
73Qwen2.5 VL 72B InstructAlibaba128K~171$0.25$0.75
74Sonar Deep ResearchPerplexity128K~171$2.00$8.00
75Sonar Reasoning ProPerplexity128K~171$2.00$8.00
76UI-TARS 7B ByteDance128K~171$0.10$0.20

32K–64K Tokens

(26 models)

Moderate context windows suitable for shorter documents, code files, and focused conversations.

#ModelProviderContext WindowPagesInput / 1MOutput / 1M
1SonarPerplexity127K~169$1.00$1.00
2ERNIE 4.5 VL 424B A47B Baidu123K~164$0.42$1.25
3Qwen3 30B A3B Thinking 2507Alibaba82K~109$0.20$2.40
4GLM 4.5VZhipu AI66K~87$0.60$1.80
5MiniMax M2-herMiniMax66K~87$0.30$1.20
6Mixtral 8x22B InstructMistral AI66K~87$2.00$6.00
7Nano Banana 2 (Gemini 3.1 Flash Image Preview)Google66K~87$0.50$3.00
8Nano Banana 2 Lite (Gemini 3.1 Flash Lite Image)Google66K~87$0.25$1.50
9Nano Banana Pro (Gemini 3 Pro Image Preview)Google66K~87$2.00$12.00
10Olmo 3 32B ThinkAllen AI66K~87$0.15$0.50
11Reka Flash 3rekaai66K~87$0.10$0.20
12WizardLM-2 8x22BMicrosoft66K~87$0.62$0.62
13Llama 3.2 1B InstructMeta60K~80$0.03$0.20
14Falcon Arabic 7B InstructTII33K~44FreeFree
15Falcon Mamba 7B InstructTII33K~44FreeFree
16Falcon3 10B InstructTII33K~44FreeFree
17Falcon3 7B InstructTII33K~44FreeFree
18Gemma 3n 4BGoogle33K~44$0.06$0.12
19Mistral Small 3Mistral AI33K~44$0.05$0.08
20Nano Banana (Gemini 2.5 Flash Image)Google33K~44$0.30$2.50
21Perceptron Mk1perceptron33K~44$0.15$1.50
22Qwen2.5 72B InstructAlibaba33K~44$0.36$0.40
23Qwen2.5 7B InstructAlibaba33K~44$0.10$0.20
24Qwen2.5 Coder 32B InstructAlibaba33K~44$0.66$1.00
25SabaMistral AI33K~44$0.20$0.60
26Voxtral Small 24B 2507Mistral AI32K~43$0.10$0.30

What Is a Context Window and Why Does It Matter?

Context window = working memory

A model's context window is the total number of tokens (roughly words) it can process in a single request. This includes both your input prompt and the model's output. A model with a 128K context window can process about 170 pages of text at once, while a 1M-token model can handle roughly 1,300 pages -- enough for entire books or large codebases.

Larger context enables new use cases

With small context windows (under 32K), you must chunk documents and use retrieval-augmented generation (RAG). Large context models eliminate this complexity for many workloads: analyzing full legal contracts, reviewing entire repositories, summarizing research paper collections, or maintaining very long conversations with full history retained.

Context size vs. effective recall

Not all context is created equal. Some models perform well on "needle in a haystack" tests at their full context length, while others degrade on information retrieval when prompts get very long. The advertised context window is the maximum, but effective performance may vary. Check our leaderboard for quality scores that account for real-world performance.

Cost implications of large context

Using a large context window means sending more tokens per request, which increases cost. For example, filling a 1M-token context at $3/1M input tokens costs $3 per request. For cost-sensitive workloads, consider whether RAG with a smaller context model might be more efficient than filling a large context window end to end.

Explore More

Dive deeper into model capabilities, compare context windows side by side, or see overall rankings across all dimensions.

Frequently Asked Questions

A large context window model can process long inputs - from 128K tokens (about 100 pages) up to 2M tokens (about 1,500 pages). This allows analyzing entire codebases, long documents, or extensive conversation histories in a single request.

Large context models excel when you need the AI to reason across an entire document at once. RAG (Retrieval-Augmented Generation) is better for searching across massive document collections. Large context is simpler to implement but costs more per request.

No. While many models advertise 1M+ token context windows, their effective recall can vary significantly. Some models lose accuracy when important information is buried in the middle of very long contexts. Our rankings account for practical performance, not just advertised limits.

Large Context AI Models - 1M+ Token (2026) | LM Market Cap