LLMs with a 1 million token context window
Qwen: Qwen3.7 Flash is the cheapest option at $0.030 per million input tokens. A million tokens is roughly 750,000 words - an entire codebase or a shelf of documents in a single prompt. These are the models that offer it, and what that window costs.
99 models qualify
Ranked by blended cost - input weighted 75%, output 25%.
| # | Model | Blended / 1M | Input / 1M | Output / 1M | Context | Capabilities |
|---|---|---|---|---|---|---|
| 1 | Auto Router (Beta) OpenRouter | $-1000000.0000 | $-1000000.0000 | $-1000000.0000 | 2M | ReasoningToolsVision |
| 2 | Fusion OpenRouter | $-1000000.0000 | $-1000000.0000 | $-1000000.0000 | 1M | |
| 3 | Pareto Code Router OpenRouter | $-1000000.0000 | $-1000000.0000 | $-1000000.0000 | 2M | |
| 4 | Auto Router OpenRouter | $-1000000.0000 | $-1000000.0000 | $-1000000.0000 | 2M | ReasoningToolsVision |
| 5 | Lyria 3 Pro Preview | Free | Free | Free | 1.0M | Free tierVision |
| 6 | Lyria 3 Clip Preview | Free | Free | Free | 1.0M | Free tierVision |
| 7 | Qwen3.7 Flash Qwen | $0.055 | $0.030 | $0.130 | 1M | ReasoningToolsVision |
| 8 | DeepSeek V4 Flash 0731 DeepSeek | $0.075 | $0.060 | $0.120 | 1.3M | ReasoningToolsOpen weights |
| 9 | DeepSeek V4 Flash 0423 DeepSeek | $0.107 | $0.086 | $0.171 | 1.0M | ReasoningToolsOpen weights |
| 10 | Laguna S 2.1 Poolside | $0.113 | $0.090 | $0.180 | 1.0M | Free tierReasoningToolsOpen weights |
| 11 | Qwen3.5-Flash Qwen | $0.114 | $0.065 | $0.260 | 1M | ReasoningToolsVision |
| 12 | $0.125 | $0.100 | $0.200 | 1.0M | ReasoningToolsVision | |
| 13 | $0.125 | $0.100 | $0.200 | 1.0M | ReasoningToolsVision | |
| 14 | Llama 4 Scout Meta Llama | $0.150 | $0.100 | $0.300 | 1.3M | ToolsVisionOpen weights |
| 15 | Gemini 2.5 Flash Lite | $0.175 | $0.100 | $0.400 | 1.0M | ReasoningToolsVision |
| 16 | GPT-4.1 Nano OpenAI | $0.175 | $0.100 | $0.400 | 1.0M | ToolsVision |
| 17 | MiMo-V2.5 Xiaomi | $0.175 | $0.140 | $0.280 | 1.1M | ReasoningToolsVisionOpen weights |
| 18 | Qwen3.8 Flash Qwen | $0.230 | $0.150 | $0.470 | 1M | ReasoningToolsVisionOpen weights |
| 19 | GLM 5.3 Flash Z.ai | $0.237 | $0.150 | $0.500 | 1.3M | ReasoningToolsVisionOpen weights |
| 20 | DeepSeek V4.1 Flash DeepSeek | $0.262 | $0.150 | $0.600 | 1.0M | ReasoningToolsVisionOpen weights |
| 21 | Llama 4 Maverick Meta Llama | $0.324 | $0.200 | $0.696 | 1.0M | ToolsVisionOpen weights |
| 22 | DeepSeek V4 Flash Vision Exp DeepSeek | $0.330 | $0.220 | $0.660 | 1.0M | ReasoningToolsVisionOpen weights |
| 23 | $0.390 | $0.195 | $0.975 | 1M | Tools | |
| 24 | Qwen Plus 0728 Qwen | $0.390 | $0.260 | $0.780 | 1M | Tools |
| 25 | Qwen-Plus Qwen | $0.390 | $0.260 | $0.780 | 1M | Tools |
Frequently asked
What is the cheapest LLM for million-token context?
Qwen: Qwen3.7 Flash at $0.030 per million input tokens and $0.130 per million output tokens, with a 1M token context window.
How many models qualify for million-token context?
99 of the 340 models tracked here meet the criteria for this list.
How much do prices vary within this category?
By roughly 1,200×, from Qwen3.7 Flash at the bottom to GPT-5.4 Pro at the top.