GPT-5.6 Luna
Fast tier of the GPT-5.6 family (GA July 9, 2026). Optimized for latency and cost while keeping strong agentic/tool-use scores (~82.5% TerminalBench). 90% cached-input discount.
GPT-5.6 Luna strengths
- Low latency and low cost
- Solid agentic tool use for its tier
- ~1M context window
- Good default for high-volume tasks
Pricing & context
| Context window | ~1M tokens (128K max output) |
| Input price /1M | $1.00 |
| Output price /1M | $6.00 |
| Modalities | text, image |
Cost guide: a typical call of about 10K input + 2K output tokens costs roughly $0.022 at list prices. Worth modelling against cheaper tiers before committing high-volume traffic.
When to choose GPT-5.6 Luna
GPT-5.6 Luna is best for high-volume, latency-sensitive workloads and cheap agent loops where flagship depth isn't required. If your workload is more cost-sensitive, weigh it against gpt-oss-120b (≈$0.03 input /1M) first.
GPT-5.6 Luna FAQ
How much does GPT-5.6 Luna cost?
GPT-5.6 Luna is priced at $1.00 per 1M input tokens and $6.00 per 1M output tokens (public API list price), with a ~1M tokens (128K max output) context window. A typical call of about 10K input and 2K output tokens costs roughly $0.022.
What is GPT-5.6 Luna best for?
GPT-5.6 Luna by OpenAI is best for high-volume, latency-sensitive workloads and cheap agent loops where flagship depth isn't required.
How does GPT-5.6 Luna pricing compare to MiMo-V2.5?
GPT-5.6 Luna input costs $1.00 per 1M tokens versus $0.105 for MiMo-V2.5, roughly 9.5x more expensive on input. Output is $6.00 vs $0.28.
Is GPT-5.6 Luna multimodal?
GPT-5.6 Luna supports text, image.
Other models
All models →| 01 | Claude Fable 5 | Anthropic | $10.00 | → |
| 02 | Claude Opus 5 | Anthropic | $5.00 | → |
| 03 | GPT-5.6 Sol | OpenAI | $5.00 | → |
| 04 | GPT-5.5 | OpenAI | $5.00 | → |