OpenAI · 400K context · updated 2026-07-28
Input · what you send
$30
per 1M tokens
Output · what it writes
$180
per 1M tokens
55,556 tokens
333,333 tokens
For context: the cheapest model tracked here is Gemini 3 Flash at $3/1M output tokens, and the most expensive is GPT-5.5 Pro at $180 — a 60× spread.Try your own budget in the calculator →
| Model | Input / 1M | Output / 1M | vs GPT-5.5 Pro (output) |
|---|---|---|---|
| GPT-5.6 Sol | $5 | $30 | 6× cheaper |
| Claude Fable 5 | $10 | $50 | 3.6× cheaper |
| GPT-5.6 Terra | $2.50 | $15 | 12× cheaper |
| GPT-5.5 | $5 | $30 | 6× cheaper |
| Claude Opus 4.8 | $5 | $25 | 7.2× cheaper |
| Claude Sonnet 5 | $2 | $10 | 18× cheaper |
| Gemini 3.1 Pro | $2 | $12 | 15× cheaper |
| Gemini 3.5 Flash | $1.50 | $9 | 20× cheaper |
| GPT-5.6 Luna | $1 | $6 | 30× cheaper |
| Claude Haiku 4.5 | $1 | $5 | 36× cheaper |
| Gemini 3 Flash | $0.50 | $3 | 60× cheaper |
GPT-5.5 Pro supports a 400K token context window. That means in a single API call, you can send up to 400K tokens of input — including your prompt, system instructions, and any documents or code you want the model to process.
In practical terms, 400K tokens is approximately:
Compare all models' context windows on the context windows page.