Longest Context Window AI Models
Ranked from provider-published pricing · Prices checked 12 August 2026
The context window is the hard ceiling on how much text a model can consider in one request: prompt, retrieved documents, conversation history and its own output all share it. This list ranks the catalogue by that ceiling, largest first.
Prices are shown beside each window so the tradeoff is visible, the largest windows are not always the most expensive models.
The ranking
top 20 of 271Gemini 1.5 Pro from Google leads this ranking at 2,000,000 tokens. The median across the 271 qualifying models is 128,000 tokens.
| # | Model | Provider | Maximum Context Window | Input /1M | Output /1M | Blended 70/30 |
|---|---|---|---|---|---|---|
| 1 | Gemini 1.5 Pro | 2,000,000 tokens | $1.25 | $5.00 | $2.38 | |
| 2 | GPT-5.6 Sol | OpenAI | 1,050,000 tokens | $5.00 | $30.00 | $12.50 |
| 3 | GPT-5.6 Terra | OpenAI | 1,050,000 tokens | $2.00 | $12.00 | $5.00 |
| 4 | GPT-5.6 Luna | OpenAI | 1,050,000 tokens | $0.200 | $1.20 | $0.500 |
| 5 | GPT-5.5 | OpenAI | 1,050,000 tokens | $5.00 | $30.00 | $12.50 |
| 6 | Gemini 3.6 Flash | 1,050,000 tokens | $1.50 | $7.50 | $3.30 | |
| 7 | Muse Spark 1.2 | Meta | 1,049,000 tokens | $1.25 | $4.25 | $2.15 |
| 8 | Kimi K3 | Moonshot AI | 1,049,000 tokens | $3.00 | $15.00 | $6.60 |
| 9 | Claude Fable 5 | Anthropic | 1,000,000 tokens | $10.00 | $50.00 | $22.00 |
| 10 | Claude Opus 5 | Anthropic | 1,000,000 tokens | $5.00 | $25.00 | $11.00 |
| 11 | Claude Sonnet 5 | Anthropic | 1,000,000 tokens | $2.00 | $10.00 | $4.40 |
| 12 | Claude Opus 4.8 | Anthropic | 1,000,000 tokens | $5.00 | $25.00 | $11.00 |
| 13 | Claude Opus 4.7 | Anthropic | 1,000,000 tokens | $5.00 | $25.00 | $11.00 |
| 14 | Claude Opus 4.6 | Anthropic | 1,000,000 tokens | $5.00 | $25.00 | $11.00 |
| 15 | Claude Sonnet 4.6 | Anthropic | 1,000,000 tokens | $3.00 | $15.00 | $6.60 |
| 16 | GPT-4.1 | OpenAI | 1,000,000 tokens | $2.00 | $8.00 | $3.80 |
| 17 | GPT-4.1 mini | OpenAI | 1,000,000 tokens | $0.400 | $1.60 | $0.760 |
| 18 | GPT-4.1 nano | OpenAI | 1,000,000 tokens | $0.100 | $0.400 | $0.190 |
| 19 | Gemini 3.5 Pro | 1,000,000 tokens | $2.00 | $12.00 | $5.00 | |
| 20 | Gemini 3.5 Flash | 1,000,000 tokens | $1.50 | $9.00 | $3.75 |
How this list is built
- Ranked by maximum context window, largest first.
- Excludes embed models.
- Includes models marked available or preview; retired and deprecated SKUs are excluded.
- Prices are the rates each provider publishes, not estimates. How the blended rate is calculated.
Frequently asked questions
What is a context window?
It is the maximum number of tokens a model can process in a single request, covering everything you send plus everything it generates. Exceeding it forces you to truncate, summarise or chunk the input. As a rough guide, one million tokens is somewhere around 750,000 words.
Do I need a 1M-token context window?
Rarely. Most production systems retrieve the handful of relevant passages instead of sending an entire corpus, which is both cheaper and usually more accurate, since models can lose track of details buried in a very long context. Large windows matter most for whole-codebase reasoning and long document analysis where chunking loses the thread.
Related rankings
Price it for your workload
Rankings are recomputed on every deploy from the catalogue. Prices checked 12 August 2026.