Context · 1M+

Cheapest 1M context LLMs

Every LLM with a 1,000,000+ token context window. Ranked by input price per 1M tokens.

Models40
Cheapest$-1000000.00
Min context1M tokens
What this page is
This page lists every priced model with a context window of at least one million tokens. 1M context unlocks whole-repo coding, book-length analysis, and massive multi-document RAG without chunking. The cost per call can be steep, so compare carefully and lean on context caching whenever possible.

1M+ context models, cheapest first.

#ModelIn $/1MOut $/1MType
1openrouter logoAuto Router$-1000000.00$-1000000.00Closed
2openrouter logoFusion$-1000000.00$-1000000.00Closed
3openrouter logoPareto Code Router$-1000000.00$-1000000.00Closed
4Google DeepMind logoLyria 3 Clip Preview$0.00$0.00Closed
5Google DeepMind logoLyria 3 Pro Preview$0.00$0.00Closed
6NVIDIA logoNemotron 3 Super (free)$0.00$0.00OSS
7NVIDIA logoNemotron 3 Ultra (free)$0.00$0.00OSS
8openrouter logoOwl Alpha$0.00$0.00Closed
9Alibaba Qwen logoQwen3 Coder 480B A35B (free)$0.00$0.00OSS
10Alibaba Qwen logoQwen3.6 Plus (free)$0.00$0.00Closed
11Alibaba Qwen logoQwen3.6 Plus Preview (free)$0.00$0.00OSS
12Alibaba Qwen logoQwen3.5-Flash$0.07$0.26OSS
13Google DeepMind logoGemini 2.0 Flash Lite$0.07$0.30Closed
14NVIDIA logoNemotron 3 Super$0.08$0.45OSS
15DeepSeek logoDeepSeek V4 Flash$0.08$0.17OSS
16Google DeepMind logoGemini 2.0 Flash$0.10$0.40Closed
17Google DeepMind logoGemini 2.5 Flash Lite$0.10$0.40Closed
18Google DeepMind logoGemini 2.5 Flash Lite Preview 09-2025$0.10$0.40Closed
19OpenAI logoGPT-4.1 Nano$0.10$0.40Closed
20Meta logoLlama 4 Scout$0.10$0.30OSS
21xiaomi logoMiMo-V2.5$0.10$0.28OSS
22Meta logoLlama 4 Maverick$0.15$0.60OSS
23Alibaba Qwen logoQwen3.6 Flash$0.19$1.13OSS
24Alibaba Qwen logoQwen3 Coder Flash$0.20$0.97OSS
25xAI logoGrok 4.1 Fast$0.20$0.50Closed
26minimax logoMiniMax-01$0.20$1.10OSS
27Alibaba Qwen logoQwen3 Coder 480B A35B$0.22$1.80OSS
28Google DeepMind logoGemini 3.1 Flash Lite$0.25$1.50Closed
29Google DeepMind logoGemini 3.1 Flash Lite Preview$0.25$1.50Closed
30Alibaba Qwen logoQwen Plus 0728$0.26$0.78OSS
31Alibaba Qwen logoQwen Plus 0728 (thinking)$0.26$0.78OSS
32Alibaba Qwen logoQwen-Plus$0.26$0.78OSS
33Alibaba Qwen logoQwen3.5 Plus 2026-02-15$0.26$1.56OSS
34Google DeepMind logoGemini 2.5 Flash$0.30$2.50Closed
35minimax logoMiniMax M3$0.30$1.20OSS
36Amazon logoNova 2 Lite$0.30$2.50Closed
37Alibaba Qwen logoQwen3.5 Plus 2026-04-20$0.30$1.80OSS
38Alibaba Qwen logoQwen3.7 Plus$0.32$1.28OSS
39Alibaba Qwen logoQwen3.6 Plus$0.33$1.95OSS
40OpenAI logoGPT-4.1 Mini$0.40$1.60Closed
Cheapest
Auto Router
$-1000000.00/M
$ per 1M input tokens
Why the gap

Premium 1M-context models pay for better accuracy at the tail of the window and faster ingestion. For research and one-shot analysis, the cheap end delivers equivalent answers on most prompts.

Most expensive
GPT-4.1 Mini
$0.40/M
$ per 1M input tokens
Gemini 2.5 Pro and Flash were first to a real 1M window. Claude Sonnet extended to 1M. Qwen3 Long is a strong open-source option. MiniMax and several Chinese labs also ship 1M+. See the table above for current live list.