GPT-5.6 Luna
GPT-5.6's fastest and most affordable tier. Built for high-volume, latency-sensitive workloads at $1/$6 per 1M tokens, with a 1.05M-token context window despite the low cost.
Specifications
| Lab | OpenAI |
|---|---|
| Context window | 1.1M tokens |
| Input price | $0.20/1M |
| Output price | $1.20/1M |
| Release | Jul 2026 |
Strengths
- Cheapest GPT-5.6 tier at $1/$6 per 1M tokens
- Fastest tier built for latency-sensitive, high-volume workloads
- 1.05M-token context despite the low price point
- Good default for cost-sensitive production traffic
Weaknesses
- Lowest-capability tier not for hard reasoning
- No Ultra mode or max reasoning-effort control
Best for
- High-volume chat and classification
- Latency-sensitive APIs
- Cost-sensitive production routing
In Depth: GPT-5.6 Luna
Summary
GPT-5.6 Luna is an AI model from OpenAI.
Released in Jul 2026. It takes text input and produces text output, with a context window of 1.1M tokens. Input pricing is $0.20 per 1M tokens and output is $1.20 per 1M tokens on OpenRouter.