The fastest, cheapest member of the GPT-5.6 family — strong capability at OpenAI's lowest GPT-5.6 price, built for high-volume and latency-sensitive work.
Luna is the entry point to the GPT-5.6 family at $0.20 / 1M input and $1.20 / 1M output — cut from $1/$6 on July 30, 2026, and now a fraction of Sol's output cost for latency-sensitive, high-throughput work.
Luna is the fast-and-affordable tier of the GPT-5.6 family, released publicly on July 9, 2026 alongside Sol and Terra at $1 input and $6 output per 1M, bringing GPT-5.6-generation capability to workloads that were previously served by mini-class models but need more headroom. OpenAI cut the price further to $0.20 / $1.20 per 1M tokens on July 30, 2026.
Luna is the model to reach for when throughput and cost dominate: classification, routing, extraction, and high-volume chat where each request is relatively simple but there are a lot of them. It doesn't carry Sol's max/ultra reasoning modes, and it trails Terra on the hardest tasks, but it's dramatically cheaper.
Luna handles everyday generation, summarization, extraction, and lightweight tool use with low latency. It's a natural fit for pipelines that fan out many small calls, and for interactive apps where responsiveness matters more than frontier reasoning. For harder problems, step up to Terra or Sol.
| Model | Input / 1M | Output / 1M | Context |
|---|---|---|---|
| GPT-5.6 Luna | $0.20 | $1.20 | 1M |
| GPT-5.6 Terra | $2 | $12 | 1M |
| GPT-5.4 nano | $0.20 | $1.25 | 400K |
| GPT-5.4 mini | $0.75 | $4.50 | 400K |
| Claude Haiku 4.5 | $1 | $5 | 200K |
| Gemini 3.5 Flash | $1.50 | $9 | 1M |
Luna is now priced right alongside GPT-5.4 nano — matching its input rate and slightly undercutting it on output — while carrying full GPT-5.6-generation capability and a 1M context window. Its closest full-size rivals are Claude Haiku 4.5 and Gemini 3.5 Flash.