SmophyAI

Published daily · Next update 02:00 UTC

Context windows

What does a full context window actually cost?

SpaceXAI: Grok 4.20 has the largest tracked context window at 2.0M tokens - filling it costs $2.50 in input tokens alone. The cheapest way to fill a large context window right now is DeepSeek: DeepSeek V4 Flash 0423, at $0.09. Window size alone doesn't tell you the real cost of using it.

Ranked by context window size
#ModelContext window$/1M inputCost to fillQuality
01SpaceXAI: Grok 4.202.0M$1.25$2.50-
02SpaceXAI: Grok 4.20 Multi-Agent2.0M$1.25$2.50-
03OpenAI: GPT-5.4 (batch)1.1M$2.50$2.6353.1
04OpenAI: GPT-5.4 Pro (batch)1.1M$30.00$31.50-
05OpenAI: GPT-5.5 (batch)1.1M$5.00$5.2556.3
06OpenAI: GPT-5.5 Pro (batch)1.1M$30.00$31.50-
07OpenAI: GPT-5.6 Luna (batch)1.1M$0.10$0.1052.3
08OpenAI: GPT-5.6 Luna Pro (batch)1.1M$0.10$0.10-
09OpenAI: GPT-5.6 Sol (batch)1.1M$5.00$5.2560.9
10OpenAI: GPT-5.6 Sol Pro (batch)1.1M$5.00$5.25-
11OpenAI: GPT-5.6 Terra (batch)1.1M$1.00$1.0556.6
12OpenAI: GPT-5.6 Terra Pro (batch)1.1M$1.00$1.05-
13Xiaomi: MiMo-V2.51.1M$0.14$0.1538
14Xiaomi: MiMo-V2.5-Pro1.1M$0.43$0.4642.9
15OpenAI GPT Latest1.1M$5.00$5.25-
16Meituan: LongCat 2.01.0M$0.30$0.31-
17DeepSeek: DeepSeek V4 Flash 04231.0M$0.09$0.0951.8
18DeepSeek: DeepSeek V4 Pro1.0M$0.43$0.4645.3
19Google: Gemini 2.5 Flash (batch)1.0M$0.30$0.31-
20Google: Gemini 2.5 Flash Lite (batch)1.0M$0.10$0.10-
21Google: Gemini 2.5 Pro (batch)1.0M$1.25$1.3125.9
22Google: Gemini 3 Flash Preview (batch)1.0M$0.50$0.52-
23Google: Gemini 3.1 Flash Lite (batch)1.0M$0.25$0.26-
24Google: Gemini 3.1 Flash Lite Preview1.0M$0.25$0.2625.6
25Google: Gemini 3.1 Pro Preview (batch)1.0M$2.00$2.1047.7
26Google: Gemini 3.1 Pro Preview Custom Tools1.0M$2.00$2.10-
27Google: Gemini 3.5 Flash (batch)1.0M$1.50$1.5752
28Google: Gemini 3.5 Flash Lite (batch)1.0M$0.30$0.3137.4
29Google: Gemini 3.6 Flash (batch)1.0M$1.50$1.5751.6
30Meta: Llama Guard 4 12B1.0M$0.18$0.19-

Cost to fill = context window (tokens) × prompt price per token. This is the cost of one request using the full window as input - real output tokens and multi-turn conversations cost more.

A big context window and a high-quality model aren't the same thing - among models with a window of 1M+ tokens, OpenAI: GPT-5.6 Sol (batch) scores highest on quality (see the Context window and Quality columns above).

Frequently asked questions

Which AI model has the largest context window?

SpaceXAI: Grok 4.20 has the largest tracked context window at 2.0M tokens.

How much does it cost to use a model's full context window?

Filling SpaceXAI: Grok 4.20's 2.0M-token context window costs $2.50 in input tokens at current pricing.