Skip to content

Large Context Window AI Models

Compare 379 AI models with context windows of 32K tokens or more. The largest models support 2M tokens -- enough to process entire codebases, books, or hundreds of documents in a single prompt. Data updated hourly.

123
1M+ Tokens
153
200K+ Tokens
77
128K Tokens
26
32K–64K Tokens
2M
Largest Context
379
Models (32K+)
487K
Average Context

1M+ Tokens

(123 models)

The largest context windows available. These models can process entire codebases, full books, or hundreds of documents in a single request.

#ModelProviderContext WindowPagesInput / 1MOutput / 1M
1Grok 4.20xAI2M~3K$1.25$2.50
2Grok 4.20 Multi-AgentxAI2M~3K$1.25$2.50
3Llama 4 ScoutMeta1.3M~2K$0.10$0.30
4GPT-5.4OpenAI1.1M~1K$2.50$15.00
5GPT-5.4 (batch)OpenAI1.1M~1K$1.25$7.50
6GPT-5.4 ProOpenAI1.1M~1K$30.00$180.00
7GPT-5.4 Pro (batch)OpenAI1.1M~1K$15.00$90.00
8GPT-5.5OpenAI1.1M~1K$5.00$30.00
9GPT-5.5 (batch)OpenAI1.1M~1K$2.50$15.00
10GPT-5.5 ProOpenAI1.1M~1K$30.00$180.00
11GPT-5.5 Pro (batch)OpenAI1.1M~1K$15.00$90.00
12GPT-5.6 LunaOpenAI1.1M~1K$0.10$0.60
13GPT-5.6 Luna (batch)OpenAI1.1M~1K$0.10$0.60
14GPT-5.6 Luna ProOpenAI1.1M~1K$0.10$0.60
15GPT-5.6 Luna Pro (batch)OpenAI1.1M~1K$0.10$0.60
16GPT-5.6 SolOpenAI1.1M~1K$5.00$30.00
17GPT-5.6 Sol (batch)OpenAI1.1M~1K$2.50$15.00
18GPT-5.6 Sol ProOpenAI1.1M~1K$5.00$30.00
19GPT-5.6 Sol Pro (batch)OpenAI1.1M~1K$2.50$15.00
20GPT-5.6 TerraOpenAI1.1M~1K$1.00$6.00
21GPT-5.6 Terra (batch)OpenAI1.1M~1K$1.00$6.00
22GPT-5.6 Terra ProOpenAI1.1M~1K$1.00$6.00
23GPT-5.6 Terra Pro (batch)OpenAI1.1M~1K$1.00$6.00
24MiMo-V2.5Xiaomi1.1M~1K$0.14$0.28
25MiMo-V2.5-ProXiaomi1.1M~1K$0.43$0.87
26OpenAI GPT Latest~openai1.1M~1K$5.00$30.00
27LongCat 2.0Meituan1.0M~1K$0.30$1.20
28DeepSeek V4 Flash 0423DeepSeek1.0M~1K$0.14$0.28
29DeepSeek V4 Flash 0731DeepSeek1.0M~1K$0.08$0.18
30DeepSeek V4 Flash Latest~deepseek1.0M~1K$0.08$0.25
31DeepSeek V4 ProDeepSeek1.0M~1K$1.17$2.34
32DeepSeek V4 Pro 0813DeepSeek1.0M~1K$0.43$0.87
33Gemini 2.5 FlashGoogle1.0M~1K$0.30$2.50
34Gemini 2.5 Flash (batch)Google1.0M~1K$0.15$1.25
35Gemini 2.5 Flash LiteGoogle1.0M~1K$0.10$0.40
36Gemini 2.5 Flash Lite (batch)Google1.0M~1K$0.05$0.20
37Gemini 2.5 ProGoogle1.0M~1K$1.25$10.00
38Gemini 2.5 Pro (batch)Google1.0M~1K$0.63$5.00
39Gemini 2.5 Pro Preview 05-06Google1.0M~1K$1.25$10.00
40Gemini 2.5 Pro Preview 06-05Google1.0M~1K$1.25$10.00
41Gemini 3 Flash PreviewGoogle1.0M~1K$0.50$3.00
42Gemini 3 Flash Preview (batch)Google1.0M~1K$0.25$1.50
43Gemini 3.1 Flash LiteGoogle1.0M~1K$0.25$1.50
44Gemini 3.1 Flash Lite (batch)Google1.0M~1K$0.13$0.75
45Gemini 3.1 Flash Lite PreviewGoogle1.0M~1K$0.25$1.50
46Gemini 3.1 Pro PreviewGoogle1.0M~1K$2.00$12.00
47Gemini 3.1 Pro Preview (batch)Google1.0M~1K$1.00$6.00
48Gemini 3.1 Pro Preview Custom ToolsGoogle1.0M~1K$2.00$12.00
49Gemini 3.5 FlashGoogle1.0M~1K$1.50$9.00
50Gemini 3.5 Flash (batch)Google1.0M~1K$0.75$4.50
51Gemini 3.5 Flash LiteGoogle1.0M~1K$0.30$2.50
52Gemini 3.5 Flash Lite (batch)Google1.0M~1K$0.15$1.25
53Gemini 3.6 FlashGoogle1.0M~1K$1.50$7.50
54Gemini 3.6 Flash (batch)Google1.0M~1K$0.75$3.75
55GLM 5.2Zhipu AI1.0M~1K$0.63$1.98
56Google Gemini Flash Latest~google1.0M~1K$1.50$7.50
57Google Gemini Pro Latest~google1.0M~1K$2.00$12.00
58Inklingthinkingmachines1.0M~1K$0.95$4.05
59Kimi K3Moonshot AI1.0M~1K$3.00$15.00
60Laguna S 2.1poolside1.0M~1K$0.09$0.18
61Llama 4 MaverickMeta1.0M~1K$0.20$0.70
62Llama Guard 4 12BMeta1.0M~1K$0.18$0.18
63Lyria 3 Clip PreviewGoogle1.0M~1KFreeFree
64Lyria 3 Pro PreviewGoogle1.0M~1KFreeFree
65MiniMax M3MiniMax1.0M~1K$0.30$1.20
66MoonshotAI Kimi Latest~moonshotai1.0M~1K$2.80$14.00
67Muse Spark 1.1meta1.0M~1K$1.25$4.25
68Muse Spark 1.2meta1.0M~1K$1.25$4.25
69Nemotron 3.5 LightningNVIDIA1.0M~1K$0.10$0.25
70GPT-4.1OpenAI1.0M~1K$2.00$8.00
71GPT-4.1 (batch)OpenAI1.0M~1K$1.00$4.00
72GPT-4.1 MiniOpenAI1.0M~1K$0.40$1.60
73GPT-4.1 Mini (batch)OpenAI1.0M~1K$0.20$0.80
74GPT-4.1 NanoOpenAI1.0M~1K$0.10$0.40
75GPT-4.1 Nano (batch)OpenAI1.0M~1K$0.05$0.20
76Palmyra X5Writer1.0M~1K$0.60$6.00
77MiniMax-01MiniMax1.0M~1K$0.20$1.10
78Anthropic Claude Sonnet Latest~anthropic1M~1K$2.00$10.00
79Claude Fable 5Anthropic1M~1K$10.00$50.00
80Claude Fable 5 (batch)Anthropic1M~1K$5.00$25.00
81Claude Fable Latest~anthropic1M~1K$10.00$50.00
82Claude Opus 4.6Anthropic1M~1K$5.00$25.00
83Claude Opus 4.6 (batch)Anthropic1M~1K$2.50$12.50
84Claude Opus 4.7Anthropic1M~1K$5.00$25.00
85Claude Opus 4.7 (batch)Anthropic1M~1K$2.50$12.50
86Claude Opus 4.7 (Fast)Anthropic1M~1K$30.00$150.00
87Claude Opus 4.8Anthropic1M~1K$5.00$25.00
88Claude Opus 4.8 (batch)Anthropic1M~1K$2.50$12.50
89Claude Opus 4.8 (Fast)Anthropic1M~1K$10.00$50.00
90Claude Opus 5Anthropic1M~1K$5.00$25.00
91Claude Opus 5 (batch)Anthropic1M~1K$2.50$12.50
92Claude Opus 5 (Fast)Anthropic1M~1K$10.00$50.00
93Claude Opus Latest~anthropic1M~1K$5.00$25.00
94Claude Sonnet 4Anthropic1M~1K$3.00$15.00
95Claude Sonnet 4.5Anthropic1M~1K$3.00$15.00
96Claude Sonnet 4.5 (batch)Anthropic1M~1K$1.50$7.50
97Claude Sonnet 4.6Anthropic1M~1K$3.00$15.00
98Claude Sonnet 4.6 (batch)Anthropic1M~1K$1.50$7.50
99Claude Sonnet 5Anthropic1M~1K$2.00$10.00
100Claude Sonnet 5 (batch)Anthropic1M~1K$1.00$5.00
101Fugu Ultrasakana1M~1K$5.00$30.00
102Grok 4.3xAI1M~1K$1.25$2.50
103MiniMax M1MiniMax1M~1K$0.55$2.20
104Nemotron 3 SuperNVIDIA1M~1K$0.08$0.40
105Nemotron 3 Ultra (free)NVIDIA1M~1KFreeFree
106Nemotron 3.5 Lightning (free)NVIDIA1M~1KFreeFree
107Nova 2 LiteAmazon1M~1K$0.30$2.50
108Nova Premier 1.0Amazon1M~1K$2.50$12.50
109Qwen Plus 0728Alibaba1M~1K$0.26$0.78
110Qwen Plus 0728 (thinking)Alibaba1M~1K$0.40$1.20
111Qwen-PlusAlibaba1M~1K$0.26$0.78
112Qwen3 Coder FlashAlibaba1M~1K$0.20$0.97
113Qwen3 Coder PlusAlibaba1M~1K$0.65$3.25
114Qwen3.5 Plus 2026-02-15Alibaba1M~1K$0.26$1.56
115Qwen3.5 Plus 2026-04-20Alibaba1M~1K$0.30$1.80
116Qwen3.5-FlashAlibaba1M~1K$0.07$0.26
117Qwen3.6 FlashAlibaba1M~1K$0.19$1.13
118Qwen3.6 PlusAlibaba1M~1K$0.33$1.95
119Qwen3.7 FlashAlibaba1M~1K$0.03$0.13
120Qwen3.7 MaxAlibaba1M~1K$1.48$4.42
121Qwen3.7 PlusAlibaba1M~1K$0.32$1.28
122Qwen3.8 2.4T A95BAlibaba1M~1K$2.00$6.00
123Qwen3.8 MaxAlibaba1M~1K$2.00$6.00

200K+ Tokens

(153 models)

Extended context models ideal for long documents, legal contracts, research papers, and multi-file code analysis.

#ModelProviderContext WindowPagesInput / 1MOutput / 1M
1Inkling (batch)thinkingmachines524K~699$0.50$2.02
2Inkling Smallthinkingmachines524K~699$0.45$1.20
3MiniMax M3 (batch)MiniMax524K~699$0.15$0.60
4Solar Pro 4Upstage524K~699$0.03$0.12
5Nemotron 3 UltraNVIDIA512K~683$0.60$3.60
6Nemotron 3 Ultra (batch)NVIDIA512K~683$0.30$1.80
7GLM 5.2 (batch)Zhipu AI512K~683$0.70$2.20
8Grok 4.5xAI500K~667$2.00$6.00
9Grok 4.6xAI500K~667$2.00$6.00
10Grok Latest~x-ai500K~667$2.00$6.00
11GPT Chat LatestOpenAI400K~533$5.00$30.00
12GPT-5OpenAI400K~533$1.25$10.00
13GPT-5 (batch)OpenAI400K~533$0.63$5.00
14GPT-5 Codex (batch)OpenAI400K~533$0.63$5.00
15GPT-5 ImageOpenAI400K~533$10.00$10.00
16GPT-5 Image MiniOpenAI400K~533$2.50$2.00
17GPT-5 MiniOpenAI400K~533$0.25$2.00
18GPT-5 Mini (batch)OpenAI400K~533$0.13$1.00
19GPT-5 NanoOpenAI400K~533$0.05$0.40
20GPT-5 Nano (batch)OpenAI400K~533$0.02$0.20
21GPT-5 ProOpenAI400K~533$15.00$120.00
22GPT-5 Pro (batch)OpenAI400K~533$7.50$60.00
23GPT-5.1OpenAI400K~533$1.25$10.00
24GPT-5.1 (batch)OpenAI400K~533$0.63$5.00
25GPT-5.1-CodexOpenAI400K~533$1.25$10.00
26GPT-5.1-Codex-MaxOpenAI400K~533$1.25$10.00
27GPT-5.1-Codex-MiniOpenAI400K~533$0.25$2.00
28GPT-5.2OpenAI400K~533$1.75$14.00
29GPT-5.2 (batch)OpenAI400K~533$0.88$7.00
30GPT-5.2 ProOpenAI400K~533$21.00$168.00
31GPT-5.2 Pro (batch)OpenAI400K~533$10.50$84.00
32GPT-5.2-CodexOpenAI400K~533$1.75$14.00
33GPT-5.3-CodexOpenAI400K~533$1.75$14.00
34GPT-5.4 MiniOpenAI400K~533$0.75$4.50
35GPT-5.4 Mini (batch)OpenAI400K~533$0.38$2.25
36GPT-5.4 NanoOpenAI400K~533$0.20$1.25
37GPT-5.4 Nano (batch)OpenAI400K~533$0.10$0.63
38OpenAI GPT Mini Latest~openai400K~533$0.75$4.50
39Nova Lite 1.0Amazon300K~400$0.06$0.24
40Nova Pro 1.0Amazon300K~400$0.80$3.20
41GPT-5.4 Image 2OpenAI272K~363$8.00$15.00
42Falcon-H1-Arabic 34B InstructTII262K~350FreeFree
43Falcon-H1-Arabic 7B InstructTII262K~350FreeFree
44Gemma 3 27BGoogle262K~350$0.08$0.45
45Gemma 4 26B A4B Google262K~350$0.12$0.40
46Gemma 4 26B A4B (free)Google262K~350FreeFree
47Gemma 4 31BGoogle262K~350$0.10$0.34
48Gemma 4 31B (free)Google262K~350FreeFree
49Hy3Tencent262K~350$0.13$0.53
50Hy3 previewTencent262K~350$0.06$0.21
51KAT-Coder-Pro V2Kuaishou262K~350$0.30$1.20
52Kimi K2 0905Moonshot AI262K~350$0.60$2.50
53Kimi K2 ThinkingMoonshot AI262K~350$0.60$2.50
54Kimi K2.5Moonshot AI262K~350$0.57$2.85
55Kimi K2.6Moonshot AI262K~350$0.65$2.72
56Kimi K2.7 CodeMoonshot AI262K~350$0.67$3.40
57Kimi K2.7 Code (batch)Moonshot AI262K~350$0.47$2.00
58Laguna S 2.1 (free)poolside262K~350FreeFree
59Laguna XS 2.1poolside262K~350$0.06$0.12
60Laguna XS 2.1 (free)poolside262K~350FreeFree
61Ling-2.6-1Tinclusionai262K~350$0.07$0.63
62Ling-2.6-flashinclusionai262K~350$0.01$0.03
63Ling-3.0-flashinclusionai262K~350$0.02$0.06
64Ministral 3 14B 2512Mistral AI262K~350$0.20$0.20
65Ministral 3 8B 2512Mistral AI262K~350$0.15$0.15
66Mistral Large 3 2512Mistral AI262K~350$0.50$1.50
67Mistral Medium 3.5Mistral AI262K~350$1.50$7.50
68Mistral Small 4Mistral AI262K~350$0.15$0.60
69Nemotron 3 Nano 30B A3BNVIDIA262K~350$0.05$0.20
70Nemotron 3 Super (free)NVIDIA262K~350FreeFree
71Qwen3 235B A22B Instruct 2507Alibaba262K~350$0.09$0.55
72Qwen3 235B A22B Thinking 2507Alibaba262K~350$0.23$2.30
73Qwen3 30B A3B Instruct 2507Alibaba262K~350$0.05$0.19
74Qwen3 Coder 30B A3B InstructAlibaba262K~350$0.07$0.28
75Qwen3 Coder 480B A35BAlibaba262K~350$0.30$1.00
76Qwen3 Coder NextAlibaba262K~350$0.12$0.80
77Qwen3 MaxAlibaba262K~350$0.78$3.90
78Qwen3 Max ThinkingAlibaba262K~350$0.78$3.90
79Qwen3 Next 80B A3B InstructAlibaba262K~350$0.10$1.10
80Qwen3 Next 80B A3B ThinkingAlibaba262K~350$0.15$1.20
81Qwen3 VL 235B A22B InstructAlibaba262K~350$0.26$1.04
82Qwen3 VL 30B A3B InstructAlibaba262K~350$0.15$0.60
83Qwen3 VL 30B A3B ThinkingAlibaba262K~350$0.20$2.40
84Qwen3 VL 8B InstructAlibaba262K~350$0.12$0.45
85Qwen3.5 397B A17BAlibaba262K~350$0.50$3.60
86Qwen3.5-122B-A10BAlibaba262K~350$0.29$2.40
87Qwen3.5-27BAlibaba262K~350$0.20$1.56
88Qwen3.5-35B-A3BAlibaba262K~350$0.25$1.25
89Qwen3.5-9BAlibaba262K~350$0.10$0.15
90Qwen3.6 27BAlibaba262K~350$0.60$3.60
91Qwen3.6 35B A3BAlibaba262K~350$0.15$1.00
92Qwen3.6 Max PreviewAlibaba262K~350$1.03$6.16
93Ring-2.6-1Tinclusionai262K~350$0.07$0.63
94Sakana Namazusakana262K~350$0.95$4.00
95Seed 1.6ByteDance262K~350$0.25$2.00
96Seed 1.6 FlashByteDance262K~350$0.07$0.30
97Seed 2.1 TurboByteDance262K~350$0.50$2.50
98Seed-2.0-CodeByteDance262K~350$0.50$3.00
99Seed-2.0-LiteByteDance262K~350$0.25$2.00
100Seed-2.0-MiniByteDance262K~350$0.10$0.40
101Step 3.5 FlashStepFun262K~350$0.10$0.30
102Step 3.7 FlashStepFun262K~350$0.20$1.15
103Trinity Large Thinkingarcee-ai262K~350$0.22$0.85
104Codestral 2508Mistral AI256K~341$0.30$0.90
105Command ACohere256K~341$2.50$10.00
106Grok Build 0.1xAI256K~341$1.00$2.00
107Jamba Large 1.7AI21 Labs256K~341$2.00$8.00
108KAT-Coder-Air V2.5Kuaishou256K~341$0.15$0.60
109KAT-Coder-Pro V2.5Kuaishou256K~341$0.74$2.96
110Mistral Small 3.2 24BMistral AI256K~341$0.09$0.25
111Nemotron 3 Nano 30B A3B (free)NVIDIA256K~341FreeFree
112Nemotron 3 Nano Omni (free)NVIDIA256K~341FreeFree
113North Mini Code (free)Cohere256K~341FreeFree
114GLM 4.6Zhipu AI205K~273$0.50$2.00
115GLM 4.7Zhipu AI205K~273$0.40$1.75
116GLM 5Zhipu AI205K~273$0.95$2.55
117GLM 5.1Zhipu AI205K~273$0.95$2.99
118MiniMax M2MiniMax205K~273$0.26$1.02
119MiniMax M2.1MiniMax205K~273$0.30$1.20
120MiniMax M2.5MiniMax205K~273$0.22$0.90
121MiniMax M2.7MiniMax205K~273$0.30$1.20
122GLM 4.7 FlashZhipu AI203K~270$0.06$0.40
123GLM 5 TurboZhipu AI203K~270$1.20$4.00
124GLM 5V TurboZhipu AI203K~270$1.20$4.00
125Anthropic Claude Haiku Latest~anthropic200K~267$1.00$5.00
126Claude 3 HaikuAnthropic200K~267$0.25$1.25
127Claude Haiku 4.5Anthropic200K~267$1.00$5.00
128Claude Haiku 4.5 (batch)Anthropic200K~267$0.50$2.50
129Claude Opus 4Anthropic200K~267$15.00$75.00
130Claude Opus 4.1Anthropic200K~267$15.00$75.00
131Claude Opus 4.1 (batch)Anthropic200K~267$7.50$37.50
132Claude Opus 4.5Anthropic200K~267$5.00$25.00
133Claude Opus 4.5 (batch)Anthropic200K~267$2.50$12.50
134Composer 2Cursor200K~267$0.50$2.50
135Composer 2 FastCursor200K~267$1.50$7.50
136o1OpenAI200K~267$15.00$60.00
137o1 (batch)OpenAI200K~267$7.50$30.00
138o1-proOpenAI200K~267$150.00$600.00
139o1-pro (batch)OpenAI200K~267$75.00$300.00
140o3OpenAI200K~267$2.00$8.00
141o3 (batch)OpenAI200K~267$1.00$4.00
142o3 MiniOpenAI200K~267$1.10$4.40
143o3 Mini (batch)OpenAI200K~267$0.55$2.20
144o3 Mini HighOpenAI200K~267$1.10$4.40
145o3 Mini High (batch)OpenAI200K~267$0.55$2.20
146o3 ProOpenAI200K~267$20.00$80.00
147o3 Pro (batch)OpenAI200K~267$10.00$40.00
148o4 MiniOpenAI200K~267$1.10$4.40
149o4 Mini (batch)OpenAI200K~267$0.55$2.20
150o4 Mini HighOpenAI200K~267$1.10$4.40
151o4 Mini High (batch)OpenAI200K~267$0.55$2.20
152Sonar ProPerplexity200K~267$3.00$15.00
153Sonar Pro SearchPerplexity200K~267$3.00$15.00

128K Tokens

(77 models)

The current standard for frontier models. Sufficient for most production use cases, including long conversations and medium-length documents.

#ModelProviderContext WindowPagesInput / 1MOutput / 1M
1DeepSeek V3DeepSeek164K~218$0.26$1.03
2DeepSeek V3 0324DeepSeek164K~218$0.27$1.12
3DeepSeek V3.1DeepSeek164K~218$0.25$0.95
4DeepSeek V3.1 TerminusDeepSeek164K~218$0.27$0.95
5DeepSeek V3.2DeepSeek164K~218$0.27$0.40
6DeepSeek V3.2 ExpDeepSeek164K~218$0.27$0.41
7R1DeepSeek164K~218$0.70$2.50
8R1 0528DeepSeek164K~218$0.50$2.15
9Aion-2.0aion-labs131K~175$0.80$1.60
10Aion-3.0aion-labs131K~175$3.00$6.00
11Aion-3.0-Miniaion-labs131K~175$0.70$1.40
12Falcon-H1-Arabic 3B InstructTII131K~175FreeFree
13Gemma 3 12BGoogle131K~175$0.05$0.15
14Gemma 3 4BGoogle131K~175$0.05$0.10
15GLM 4.5Zhipu AI131K~175$0.60$2.20
16GLM 4.5 AirZhipu AI131K~175$0.13$0.85
17GLM 4.6VZhipu AI131K~175$0.30$0.90
18gpt-oss-120bOpenAI131K~175$0.03$0.17
19gpt-oss-20bOpenAI131K~175$0.03$0.13
20gpt-oss-20b (free)OpenAI131K~175FreeFree
21gpt-oss-safeguard-20bOpenAI131K~175$0.07$0.30
22Granite 4.1 8BIBM131K~175$0.05$0.10
23Hunyuan A13B InstructTencent131K~175$0.14$0.57
24Kimi K2 0711Moonshot AI131K~175$0.57$2.30
25Llama 3.1 70B InstructMeta131K~175$0.40$0.40
26Llama 3.1 8B InstructMeta131K~175$0.05$0.08
27Llama 3.2 3B InstructMeta131K~175$0.05$0.33
28Llama 3.3 70B InstructMeta131K~175$0.10$0.32
29Ministral 3 3B 2512Mistral AI131K~175$0.10$0.10
30Mistral Large 2407Mistral AI131K~175$2.00$6.00
31Mistral Medium 3Mistral AI131K~175$0.40$2.00
32Mistral Medium 3.1Mistral AI131K~175$0.40$2.00
33Mistral NemoMistral AI131K~175$0.02$0.03
34Muse Glimmer 30Bmeta131K~175$0.35$1.50
35Nano Banana 2 (Gemini 3.1 Flash Image)Google131K~175$0.50$3.00
36Nano Banana Pro (Gemini 3 Pro Image)Google131K~175$2.00$12.00
37Qwen3 14BAlibaba131K~175$0.12$0.24
38Qwen3 235B A22BAlibaba131K~175$0.45$1.82
39Qwen3 30B A3BAlibaba131K~175$0.12$0.50
40Qwen3 32BAlibaba131K~175$0.08$0.28
41Qwen3 8BAlibaba131K~175$0.12$0.45
42Qwen3 VL 235B A22B ThinkingAlibaba131K~175$0.40$4.00
43Qwen3 VL 32B InstructAlibaba131K~175$0.10$0.42
44Qwen3 VL 8B ThinkingAlibaba131K~175$0.18$2.10
45Solar Pro 3Upstage131K~175$0.15$0.60
46Virtuoso Largearcee-ai131K~175$0.75$1.20
47Granite 4.0 MicroIBM131K~175$0.02$0.11
48Cogito v2.1 671Bdeepcogito128K~171$1.25$1.25
49Command R (08-2024)Cohere128K~171$0.15$0.60
50Command R+ (08-2024)Cohere128K~171$2.50$10.00
51Command R7B (12-2024)Cohere128K~171$0.04$0.15
52GPT AudioOpenAI128K~171$2.50$10.00
53GPT Audio MiniOpenAI128K~171$0.60$2.40
54GPT-4 TurboOpenAI128K~171$10.00$30.00
55GPT-4 Turbo (batch)OpenAI128K~171$5.00$15.00
56GPT-4 Turbo PreviewOpenAI128K~171$10.00$30.00
57GPT-4oOpenAI128K~171$2.50$10.00
58GPT-4o (2024-05-13)OpenAI128K~171$5.00$15.00
59GPT-4o (2024-08-06)OpenAI128K~171$2.50$10.00
60GPT-4o (2024-11-20)OpenAI128K~171$2.50$10.00
61GPT-4o (batch)OpenAI128K~171$1.25$5.00
62GPT-4o-miniOpenAI128K~171$0.15$0.60
63GPT-4o-mini (2024-07-18)OpenAI128K~171$0.15$0.60
64GPT-4o-mini (batch)OpenAI128K~171$0.07$0.30
65GPT-5.2 ChatOpenAI128K~171$1.75$14.00
66LFM2.5-2.6B (free)Liquid AI128K~171FreeFree
67Mercury 2Inception128K~171$0.25$0.75
68Mistral LargeMistral AI128K~171$2.00$6.00
69Mistral Small 3.1 24BMistral AI128K~171$0.35$0.55
70Nemotron 3.5 Content Safety (free)NVIDIA128K~171FreeFree
71Nemotron Nano 12B 2 VL (free)NVIDIA128K~171FreeFree
72Nemotron Nano 9B V2 (free)NVIDIA128K~171FreeFree
73Nova Micro 1.0Amazon128K~171$0.04$0.14
74Qwen2.5 VL 72B InstructAlibaba128K~171$0.25$0.75
75Sonar Deep ResearchPerplexity128K~171$2.00$8.00
76Sonar Reasoning ProPerplexity128K~171$2.00$8.00
77UI-TARS 7B ByteDance128K~171$0.10$0.20

32K–64K Tokens

(26 models)

Moderate context windows suitable for shorter documents, code files, and focused conversations.

#ModelProviderContext WindowPagesInput / 1MOutput / 1M
1SonarPerplexity127K~169$1.00$1.00
2ERNIE 4.5 VL 424B A47B Baidu123K~164$0.42$1.25
3Qwen3 30B A3B Thinking 2507Alibaba82K~109$0.20$2.40
4GLM 4.5VZhipu AI66K~87$0.60$1.80
5MiniMax M2-herMiniMax66K~87$0.30$1.20
6Mixtral 8x22B InstructMistral AI66K~87$2.00$6.00
7Nano Banana 2 (Gemini 3.1 Flash Image Preview)Google66K~87$0.50$3.00
8Nano Banana 2 Lite (Gemini 3.1 Flash Lite Image)Google66K~87$0.25$1.50
9Nano Banana Pro (Gemini 3 Pro Image Preview)Google66K~87$2.00$12.00
10Olmo 3 32B ThinkAllen AI66K~87$0.15$0.50
11Reka Flash 3rekaai66K~87$0.10$0.20
12WizardLM-2 8x22BMicrosoft66K~87$0.62$0.62
13Llama 3.2 1B InstructMeta60K~80$0.03$0.20
14Falcon Arabic 7B InstructTII33K~44FreeFree
15Falcon Mamba 7B InstructTII33K~44FreeFree
16Falcon3 10B InstructTII33K~44FreeFree
17Falcon3 7B InstructTII33K~44FreeFree
18Gemma 3n 4BGoogle33K~44$0.06$0.12
19Mistral Small 3Mistral AI33K~44$0.05$0.08
20Nano Banana (Gemini 2.5 Flash Image)Google33K~44$0.30$2.50
21Perceptron Mk1perceptron33K~44$0.15$1.50
22Qwen2.5 72B InstructAlibaba33K~44$0.36$0.40
23Qwen2.5 7B InstructAlibaba33K~44$0.10$0.20
24Qwen2.5 Coder 32B InstructAlibaba33K~44$0.66$1.00
25SabaMistral AI33K~44$0.20$0.60
26Voxtral Small 24B 2507Mistral AI32K~43$0.10$0.30

What Is a Context Window and Why Does It Matter?

Context window = working memory

A model's context window is the total number of tokens (roughly words) it can process in a single request. This includes both your input prompt and the model's output. A model with a 128K context window can process about 170 pages of text at once, while a 1M-token model can handle roughly 1,300 pages -- enough for entire books or large codebases.

Larger context enables new use cases

With small context windows (under 32K), you must chunk documents and use retrieval-augmented generation (RAG). Large context models eliminate this complexity for many workloads: analyzing full legal contracts, reviewing entire repositories, summarizing research paper collections, or maintaining very long conversations with full history retained.

Context size vs. effective recall

Not all context is created equal. Some models perform well on "needle in a haystack" tests at their full context length, while others degrade on information retrieval when prompts get very long. The advertised context window is the maximum, but effective performance may vary. Check our leaderboard for quality scores that account for real-world performance.

Cost implications of large context

Using a large context window means sending more tokens per request, which increases cost. For example, filling a 1M-token context at $3/1M input tokens costs $3 per request. For cost-sensitive workloads, consider whether RAG with a smaller context model might be more efficient than filling a large context window end to end.

探索更多

Dive deeper into model capabilities, compare context windows side by side, or see overall rankings across all dimensions.

Frequently Asked Questions

A large context window model can process long inputs - from 128K tokens (about 100 pages) up to 2M tokens (about 1,500 pages). This allows analyzing entire codebases, long documents, or extensive conversation histories in a single request.

Large context models excel when you need the AI to reason across an entire document at once. RAG (Retrieval-Augmented Generation) is better for searching across massive document collections. Large context is simpler to implement but costs more per request.

No. While many models advertise 1M+ token context windows, their effective recall can vary significantly. Some models lose accuracy when important information is buried in the middle of very long contexts. Our rankings account for practical performance, not just advertised limits.

Large Context AI Models - 1M+ Token (2026) | LM Market Cap