- Rankings
- AI models with the longest context windows
AI models with the longest context windows
Ranked by maximum context, with the price to fill it. A million-token window costs whatever a million input tokens cost, so the useful comparison is window size and input rate together — which is what this table shows.
List prices checked
All 25, ranked
| # | Name | Provider | Input | Notes |
|---|---|---|---|---|
| 1 | Gemini 2.5 Pro | $1.25/1M tokens | 1M context · $1.25 in / $10.00 out | |
| 2 | Gemini 2.5 Flash | $0.300/1M tokens | 1M context · $0.300 in / $2.50 out | |
| 3 | Gemini 2.5 Flash-Lite | $0.100/1M tokens | 1M context · $0.100 in / $0.400 out | |
| 4 | GPT-4.1 | OpenAI | $2.00/1M tokens | 1M context · $2.00 in / $8.00 out |
| 5 | GPT-4.1 mini | OpenAI | $0.400/1M tokens | 1M context · $0.400 in / $1.60 out |
| 6 | DeepSeek-V4-Flash | DeepSeek | $0.140/1M tokens | 1M context · $0.140 in / $0.280 out |
| 7 | DeepSeek-V4-Pro | DeepSeek | $0.435/1M tokens | 1M context · $0.435 in / $0.870 out |
| 8 | GPT-5 | OpenAI | $1.25/1M tokens | 400K context · $1.25 in / $10.00 out |
| 9 | GPT-5 mini | OpenAI | $0.250/1M tokens | 400K context · $0.250 in / $2.00 out |
| 10 | GPT-5 nano | OpenAI | $0.050/1M tokens | 400K context · $0.050 in / $0.400 out |
| 11 | Codestral | Mistral AI | $0.300/1M tokens | 262K context · $0.300 in / $0.900 out |
| 12 | Grok 4 | xAI | $3.00/1M tokens | 256K context · $3.00 in / $15.00 out |
| 13 | Cohere Command A | Cohere | $2.50/1M tokens | 256K context · $2.50 in / $10.00 out |
| 14 | Claude Fable 5 | Anthropic | $10.00/1M tokens | 200K context · $10.00 in / $50.00 out |
| 15 | Claude Opus 5 | Anthropic | $5.00/1M tokens | 200K context · $5.00 in / $25.00 out |
| 16 | Claude Sonnet 5 | Anthropic | $3.00/1M tokens | 200K context · $3.00 in / $15.00 out |
| 17 | Claude Haiku 4.5 | Anthropic | $1.00/1M tokens | 200K context · $1.00 in / $5.00 out |
| 18 | OpenAI o3 | OpenAI | $2.00/1M tokens | 200K context · $2.00 in / $8.00 out |
| 19 | OpenAI o4-mini | OpenAI | $1.10/1M tokens | 200K context · $1.10 in / $4.40 out |
| 20 | Grok 3 mini | xAI | $0.300/1M tokens | 131K context · $0.300 in / $0.500 out |
| 21 | Mistral Large | Mistral AI | $2.00/1M tokens | 131K context · $2.00 in / $6.00 out |
| 22 | Mistral Small | Mistral AI | $0.100/1M tokens | 131K context · $0.100 in / $0.300 out |
| 23 | GPT-4o | OpenAI | $2.50/1M tokens | 128K context · $2.50 in / $10.00 out |
| 24 | GPT-4o mini | OpenAI | $0.150/1M tokens | 128K context · $0.150 in / $0.600 out |
| 25 | Cohere Command R | Cohere | $0.150/1M tokens | 128K context · $0.150 in / $0.600 out |
Source: Google (May 2026) · OpenAI (May 2026) · DeepSeek (August 2026) · Mistral AI (May 2026) · xAI (May 2026) · Cohere (May 2026). Published list prices only — committed-use discounts, regional variation, and per-request charges are not included. Not yet independently re-checked against the vendor’s page.
Questions
- What is the ai models with the longest context windows in 2026?
- Gemini 2.5 Pro from Google tops this ranking at $1.25/1M tokens. The full table below shows all 25 entries with the same measure applied to each.
- How is this ranking calculated?
- Ranked by maximum context, with the price to fill it. A million-token window costs whatever a million input tokens cost, so the useful comparison is window size and input rate together — which is what this table shows. Every figure is the provider's own published list price, linked in the source column and stamped with the month it was checked.
- Why is the cheapest option not always the right one?
- Because a single axis never captures the whole bill. A provider with the lowest storage rate may have the highest egress rate; a model with the cheapest input tokens may be the most expensive for agentic work. The comparison pages model both sides against identical workloads, which is the check worth doing before switching.
Other rankings
- Cheapest CDN providers by egress price
- Cheapest object storage per GB-month
- Cheapest VPS and cloud servers
- Cheapest cloud egress rates, all categories
- Cloud services with no egress fees
- Cheapest managed database hosting
- Cheapest application hosting platforms
- Cloud services with the best free tiers
- Cheapest AI model APIs by blended token price
- Cheapest frontier AI models
- Cheapest AI models by output token price