GPT-5.6 Luna

by OpenAI Fast Intelligent Coding

GPT-5.6's fastest and most affordable tier. Built for high-volume, latency-sensitive workloads at $1/$6 per 1M tokens, with a 1.05M-token context window despite the low cost.

Choose a model to compare against GPT-5.6 Luna

Specifications

Specifications for GPT-5.6 Luna
LabOpenAI
Context window1.1M tokens
Input price $0.20/1M
Output price $1.20/1M
ReleaseJul 2026

Strengths

  • Cheapest GPT-5.6 tier at $1/$6 per 1M tokens
  • Fastest tier built for latency-sensitive, high-volume workloads
  • 1.05M-token context despite the low price point
  • Good default for cost-sensitive production traffic

Weaknesses

  • Lowest-capability tier not for hard reasoning
  • No Ultra mode or max reasoning-effort control

Best for

  • High-volume chat and classification
  • Latency-sensitive APIs
  • Cost-sensitive production routing

In Depth: GPT-5.6 Luna

Summary

GPT-5.6 Luna is an AI model from OpenAI.

Released in Jul 2026. It takes text input and produces text output, with a context window of 1.1M tokens. Input pricing is $0.20 per 1M tokens and output is $1.20 per 1M tokens on OpenRouter.

Sources & Further Reading

Related Models