GPT-5.6 Luna Pro pricing & benchmarks
GPT-5.6 Luna Pro is available from OpenAI at $0.100 per million input tokens and $0.600 per million output tokens ($0.225 blended at 3:1). It accepts up to 1,050,000 tokens of context. GPT-5.6 Luna Pro is the same underlying model as [GPT-5.6 Luna](https://openrouter.ai/openai/gpt-5.6-luna), served with `reasoning.mode` set to `pro` for higher-quality responses on complex tasks. Learn more in OpenAI's docs: https://devel…
Pricing
Per million tokens · OpenRouter
Input$0.100
Output$0.600
Blended (3:1)$0.225
Specification
openai/gpt-5.6-luna-pro
Context window1.1M
Max output128K
ReleasedJul 2026
LicenceProprietary
Modalitiesin:text, in:image, in:file, out:text
No independently-run benchmark scores are published for GPT-5.6 Luna Pro yet. We show nothing rather than reprinting an unverified figure.
Where to run GPT-5.6 Luna Pro
5 options across 2 providers · 22× between cheapest and dearest. Blended at 3:1. ⚠ marks an option serving less context than the best available, or re-pricing above a prompt-length threshold. Bracketed labels are the provider's own service tier or region.
| Host | Blended $/1M | In / Out | Context | Uptime 24h |
|---|---|---|---|---|
| OpenAI (flex) ⚠ re-prices above 272K prompt tokens ($0.100/$0.450) |
$0.112 | $0.050 / $0.300 | 1.1M | 99.8% |
| OpenAI ⚠ re-prices above 272K prompt tokens ($0.200/$0.900) |
$0.225 | $0.100 / $0.600 | 1.1M | 99.8% |
| OpenAI (priority) | $0.450 | $0.200 / $1.20 | 1.1M | 99.8% |
| Azure ⚠ re-prices above 272K prompt tokens ($2.00/$9.00) |
$2.25 | $1.00 / $6.00 | 1.1M | 99.6% |
| Azure (eu) ⚠ re-prices above 272K prompt tokens ($2.20/$9.90) |
$2.48 | $1.10 / $6.60 | 1.1M | — |
Host prices and uptime from OpenRouter, synced 2026-08-08.
Compare GPT-5.6 Luna Pro
FAQ
How much does GPT-5.6 Luna Pro cost?
GPT-5.6 Luna Pro costs $0.100 per million input tokens and $0.600 per million output tokens on OpenRouter. At a 3:1 input:output mix that blends to $0.225 per million tokens.
What is GPT-5.6 Luna Pro's context window?
GPT-5.6 Luna Pro accepts up to 1,050,000 tokens of context, and can generate up to 128,000 output tokens.