GPT-5.6 Luna
The fastest, cheapest member of the GPT-5.6 family and the site of the steepest price cut of the year from a US lab: on 30 Jul 2026 OpenAI dropped Luna 80%, from $1/$6 to $0.20 input / $1.20 output per 1M tokens. Cache reads run $0.02. That is an explicit answer to DeepSeek and the Chinese open-weight tier on price — though at $0.20/$1.20 Luna is still comfortably more expensive than DeepSeek V4 Flash at $0.14/$0.28. Full 1.05M context, unlike most rivals' cheap tiers.
1M
tokens
128K
tokens
$0.2
per 1M tokens
$1.2
per 1M tokens
186
tokens/sec
Jul 2026
2026-07-09
$0.70
per 1M tokens
118.6
quality per $
Capabilities
Benchmarks
Pricing History
About GPT-5.6 Luna
GPT-5.6 Luna is a fast AI model by OpenAI, released on July 9, 2026. It supports a context window of 1050K tokens and can generate up to 128K output tokens.
At $0.2 per million input tokens and $1.2 per million output tokens, its blended cost of $0.70/1M tokens makes it one of the most affordable models available. Its value score of 118.6 reflects the balance of quality and cost.
Using GPT-5.6 Luna with Swfte
Access GPT-5.6 Luna through Swfte Connect, our unified LLM gateway. Connect gives you a single API for 50+ models, with automatic routing, cost optimization, and fallback handling. You can also try GPT-5.6 Luna in our AI Playground before integrating.