Context · 1M+

Cheapest 1M context LLMs

Every LLM with a 1,000,000+ token context window. Ranked by input price per 1M tokens.

Models40
Cheapest$0.00
Min context1M tokens
What this page is
This page lists every priced model with a context window of at least one million tokens. 1M context unlocks whole-repo coding, book-length analysis, and massive multi-document RAG without chunking. The cost per call can be steep, so compare carefully and lean on context caching whenever possible.

1M+ context models, cheapest first.

#ModelIn $/1MOut $/1MType
1DeepSeek logoDeepSeek V4 Flash 0731 (free)$0.00$0.00Closed
2Google DeepMind logoLyria 3 Clip Preview$0.00$0.00Closed
3Google DeepMind logoLyria 3 Pro Preview$0.00$0.00Closed
4minimax logoMiniMax M3 (free)$0.00$0.00Closed
5NVIDIA logoNemotron 3 Ultra (free)$0.00$0.00OSS
6NVIDIA logoNemotron 3.5 Lightning (free)$0.00$0.00Closed
7openrouter logoOwl Alpha$0.00$0.00Closed
8Alibaba Qwen logoQwen3 Coder 480B A35B (free)$0.00$0.00OSS
9Alibaba Qwen logoQwen3.6 Plus (free)$0.00$0.00Closed
10Alibaba Qwen logoQwen3.6 Plus Preview (free)$0.00$0.00OSS
11DeepSeek logoDeepSeek V4 Flash 0731$0.02$1.28Closed
12DeepSeek logoDeepSeek V4 Flash$0.03$1.28OSS
13Alibaba Qwen logoQwen3.7 Flash$0.03$0.13Closed
14OpenAI logoGPT-6 Luna (batch)$0.05$0.25Closed
15OpenAI logoGPT-6 Luna Pro (batch)$0.05$0.25Closed
16Google DeepMind logoGemini 2.5 Flash Lite (batch)$0.05$0.20Closed
17OpenAI logoGPT-4.1 Nano (batch)$0.05$0.20Closed
18z-ai logoGLM 5.3 Flash (batch)$0.06$0.20Closed
19Alibaba Qwen logoQwen3.5-Flash$0.07$0.26OSS
20Google DeepMind logoGemini 2.0 Flash Lite$0.07$0.30Closed
21OpenAI logoGPT-6 Luna$0.10$0.50Closed
22Google DeepMind logoGemini 2.0 Flash$0.10$0.40Closed
23Google DeepMind logoGemini 2.5 Flash Lite$0.10$0.40Closed
24Google DeepMind logoGemini 2.5 Flash Lite Preview 09-2025$0.10$0.40Closed
25OpenAI logoGPT-4.1 Nano$0.10$0.40Closed
26OpenAI logoGPT-5.6 Luna (batch)$0.10$0.60Closed
27OpenAI logoGPT-5.6 Luna Pro (batch)$0.10$0.60Closed
28OpenAI logoGPT-6 Luna Pro$0.10$0.50Closed
29Meta logoLlama 4 Scout$0.10$0.30OSS
30DeepSeek logoDeepSeek V4 Flash Vision Exp (batch)$0.11$0.33Closed
31DeepSeek logoDeepSeek V4.1 Flash (batch)$0.11$0.34Closed
32Google DeepMind logoGemini 3.1 Flash Lite (batch)$0.13$0.75Closed
33DeepSeek logoDeepSeek V4 Flash 0731 (batch)$0.14$0.28Closed
34xiaomi logoMiMo-V2.5$0.14$0.28OSS
35xiaomi logoMiMo-V2.6-Flash$0.14$0.28Closed
36DeepSeek logoDeepSeek V4.1 Flash$0.15$0.60Closed
37Google DeepMind logoGemini 2.5 Flash (batch)$0.15$1.25Closed
38Google DeepMind logoGemini 3.5 Flash Lite (batch)$0.15$1.25Closed
39z-ai logoGLM 5.3 Flash$0.15$0.50Closed
40Alibaba Qwen logoQwen3.8 Flash$0.15$0.47Closed
Cheapest
DeepSeek V4 Flash 0731 (free)
$0.00/M
$ per 1M input tokens
Why the gap

Premium 1M-context models pay for better accuracy at the tail of the window and faster ingestion. For research and one-shot analysis, the cheap end delivers equivalent answers on most prompts.

Most expensive
DeepSeek V4.1 Flash
$0.15/M
$ per 1M input tokens
Gemini 2.5 Pro and Flash were first to a real 1M window. Claude Sonnet extended to 1M. Qwen3 Long is a strong open-source option. MiniMax and several Chinese labs also ship 1M+. See the table above for current live list.