CONTEXT WINDOW COMPARISON

How much text fits in each AI model.

Every model's context window, sorted largest first, with token limits translated into approximate English words and pages. Click any model to open its token counter and check whether your actual text fits. Data refreshed 2026-08-11.

ModelProviderContext window (tokens)~Words~Pages
Llama 4 ScoutMeta10,000,000~7,500,000~15,000
Gemini 3.6 FlashGoogle1,048,576~786,000~1,573
Gemini 3.5 Flash-LiteGoogle1,048,576~786,000~1,573
Gemini 3.5 FlashGoogle1,048,576~786,000~1,573
Gemini 3.1 Pro PreviewGoogle1,048,576~786,000~1,573
Gemini 3.1 Flash-LiteGoogle1,048,576~786,000~1,573
Gemini 2.5 ProGoogle1,048,576~786,000~1,573
Gemini 2.5 FlashGoogle1,048,576~786,000~1,573
Gemini 2.5 Flash-LiteGoogle1,048,576~786,000~1,573
GPT-5.6 SolOpenAI1,047,576~786,000~1,571
GPT-5.6 TerraOpenAI1,047,576~786,000~1,571
GPT-5.6 LunaOpenAI1,047,576~786,000~1,571
GPT-5.5OpenAI1,047,576~786,000~1,571
GPT-5.5 ProOpenAI1,047,576~786,000~1,571
GPT-5.4OpenAI1,047,576~786,000~1,571
GPT-5.4 ProOpenAI1,047,576~786,000~1,571
GPT-4.1OpenAI1,047,576~786,000~1,571
GPT-4.1 miniOpenAI1,047,576~786,000~1,571
GPT-4.1 nanoOpenAI1,047,576~786,000~1,571
Claude Fable 5Anthropic1,000,000~750,000~1,500
Claude Opus 5Anthropic1,000,000~750,000~1,500
Claude Sonnet 5Anthropic1,000,000~750,000~1,500
Claude Opus 4.8Anthropic1,000,000~750,000~1,500
Claude Sonnet 4.6Anthropic1,000,000~750,000~1,500
DeepSeek V4 FlashDeepSeek1,000,000~750,000~1,500
DeepSeek V4 ProDeepSeek1,000,000~750,000~1,500
Grok 4.3xAI1,000,000~750,000~1,500
Qwen3.7 MaxAlibaba1,000,000~750,000~1,500
Qwen3.5 PlusAlibaba1,000,000~750,000~1,500
Llama 4 MaverickMeta1,000,000~750,000~1,500
Grok 4.5xAI500,000~375,000~750
GPT-5.4 miniOpenAI400,000~300,000~600
GPT-5.4 nanoOpenAI400,000~300,000~600
GPT-5.2OpenAI400,000~300,000~600
GPT-5.2 ProOpenAI400,000~300,000~600
GPT-5.1OpenAI400,000~300,000~600
GPT-5OpenAI400,000~300,000~600
GPT-5 miniOpenAI400,000~300,000~600
GPT-5 nanoOpenAI400,000~300,000~600
GPT-5 ProOpenAI400,000~300,000~600
Mistral Medium 3.5Mistral262,144~197,000~393
Mistral Large 3Mistral262,144~197,000~393
Kimi K2.6Moonshot AI262,144~197,000~393
Kimi K2 ThinkingMoonshot AI262,144~197,000~393
Grok Code Fast 1xAI256,000~192,000~384
Grok Build 0.1xAI256,000~192,000~384
Qwen3 MaxAlibaba252,000~189,000~378
o3OpenAI200,000~150,000~300
o4-miniOpenAI200,000~150,000~300
Claude Sonnet 4.5Anthropic200,000~150,000~300
Claude Haiku 4.5Anthropic200,000~150,000~300
GPT-4oOpenAI128,000~96,000~192
GPT-4o miniOpenAI128,000~96,000~192
Magistral MediumMistral128,000~96,000~192

Words and pages use English-prose averages (0.75 words per token, 500 words per page); code and non-English text fit less. Check whether a specific document fits with the words to tokens converter, or paste it into the token counter for a live context bar.

FAQ

Context window FAQ

What is a context window?

The context window is the maximum number of tokens a model can process in one request — the system prompt, conversation history, retrieved documents, and the response all share it. Text beyond the window is truncated or rejected, so long-document workflows are constrained by this limit.

Which AI model has the largest context window?

In this catalog, Llama 4 Scout has the largest context window at 10,000,000 tokens — roughly 7,500,000 English words. Several other models support 1 million tokens, while most flagship models offer 200,000 to 400,000.

How many pages fit in a 200K context window?

A 200,000-token window holds roughly 150,000 English words, or about 300 pages at 500 words per page. In practice you should budget less, because the response and system instructions consume the same window.

Does a bigger context window cost more to use?

The window itself is free — you pay for the tokens you actually send. But large windows invite large prompts, and some providers bill higher per-token rates above a threshold (for example, beyond 200K tokens on certain Gemini models). Filling a 1M window can cost dollars per request at flagship rates.