OpenAI · Efficient
GPT-5.6 Luna
The cheap tier of the GPT-5.6 family. At $0.20 per million input tokens it is priced to compete with the open-weight hosted models rather than with frontier tiers.
What that means
A worked example
Per-token pricing is hard to feel. Take a moderately chatty production workload: 100 calls a day, roughly 10,000 tokens in and 2,000 out each time. On GPT-5.6 Luna that comes to about:
$13 / month
Before caching, batching or volume discounts, all of which move this number a long way. Illustrative only.
Strong at
- -Very low cost
- -High throughput
Typical use
- -Classification
- -Bulk processing
- -Simple chat
Same tier
What else to look at
| Model | Developer | Context | In / 1M | Out / 1M |
|---|---|---|---|---|
| Claude Haiku 4.5 | Anthropic | 200K | $1 | $5 |
| DeepSeek V4 Pro | DeepSeek | 1M | $0.43 | $0.87 |
| DeepSeek V4 Flash | DeepSeek | 1M | $0.14 | $0.28 |
| Qwen 3.7 Flash | Alibaba | - | $0.03 | $0.13 |
| Qwen 3.6 27B | Alibaba | - | - | - |
Verified 2026-08-07.