All models
GPT-5.6 Luna
openai/gpt-5.6-luna
GPT-5.6 Luna is a fast, cost-efficient model in OpenAI's GPT-5.6 series. It is suited for high-volume, latency-sensitive tasks such as chat, classification, and lightweight agentic workflows, providing capable reasoning for...
Modalities
File
Image
Text
In / Out price
$1.00 / $6.00 per 1M
Context
1.05M · 128K out
Released
Jul 9, 2026
Providers
Live list prices and uptime across 5 providers serving this model.
| Provider | Input /M | Output /M | Cache read /M | Context | Max output | Uptime 24h |
|---|---|---|---|---|---|---|
| OpenAI | $0.50 | $3.00 | $0.05 | 1.05M | 128K | 98.01% |
| OpenAI | $0.25 | $1.50 | $0.025 | 1.05M | 128K | 98.01% |
| OpenAI | $1.00 | $6.00 | $0.10 | 1.05M | 128K | 98.01% |
| Azure | $1.00 | $6.00 | $0.10 | 1.05M | 128K | 99.96% |
| Azure | $1.10 | $6.60 | $0.11 | 1.05M | 128K | 100.00% |
No implicit prompt caching; cache reads require explicit cache control.
Capabilities
Tool calling
Reasoning
Moderated
include_reasoning
max_completion_tokens
max_tokens
reasoning
reasoning_effort
response_format
seed
structured_outputs
tool_choice
Pricing detail
- Input
- $1.00
- Output
- $6.00
- Cache read
- $0.10
- Cache write
- $1.25
USD per 1M tokens.