This model was retired on 2028-01-11. The pricing below is the last-known rate, kept for migration reference.
GPT-5.6 LunaDeprecated
VisionTool useReasoningPrompt cachingStructured outputWeb search
Context
922K
Max output
128K
Serving providers
3
Cheapest input
$0.20 /1M
Prices are on-demand list rates. Enterprise commitments (committed-use discounts, provisioned throughput, negotiated rates) can be materially lower.
Benchmarks
independent evaluations · source: ARC Prize and Artificial AnalysisSpeed & latency
Output speed
127 t/s
tokens / second
Time to first token
138.71s
latency
Headline indices
Intelligence Index
52.3
of 100
Coding Index
71.4
of 100
Agentic Index
46.9
of 100
Individual benchmarks · click a row for effort variants
| Benchmark | Best score | Effort | Cost / task | |||||||||||||||||||||||||||||
|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|
| APEX Agents | 36% | default | — | |||||||||||||||||||||||||||||
| ARC-AGI-2 | 60% | max | $0.670 | |||||||||||||||||||||||||||||
| ||||||||||||||||||||||||||||||||
| Automation Bench | 0.4 | default | — | |||||||||||||||||||||||||||||
| CritPt | 21% | xhigh | — | |||||||||||||||||||||||||||||
| ||||||||||||||||||||||||||||||||
| GDPval | 1578.3 | default | — | |||||||||||||||||||||||||||||
| ||||||||||||||||||||||||||||||||
| GPQA Diamond | 91% | default | — | |||||||||||||||||||||||||||||
| ||||||||||||||||||||||||||||||||
| Harvey Lab | 0.9 | default | — | |||||||||||||||||||||||||||||
| Humanity's Last Exam | 39% | default | — | |||||||||||||||||||||||||||||
| ||||||||||||||||||||||||||||||||
| IT-Bench SRE | 40% | default | — | |||||||||||||||||||||||||||||
| Long-Context Reasoning | 78% | default | — | |||||||||||||||||||||||||||||
| ||||||||||||||||||||||||||||||||
| Mlcr Overall | 0.2 | default | — | |||||||||||||||||||||||||||||
| ||||||||||||||||||||||||||||||||
| MMMU-Pro | 79% | default | — | |||||||||||||||||||||||||||||
| ||||||||||||||||||||||||||||||||
| Omniscience | -10.3 | default | — | |||||||||||||||||||||||||||||
| ||||||||||||||||||||||||||||||||
| Omniscience Accuracy | 0.4 | default | — | |||||||||||||||||||||||||||||
| ||||||||||||||||||||||||||||||||
| Omniscience Non Hallucination | 0.3 | reasoning: false | — | |||||||||||||||||||||||||||||
| ||||||||||||||||||||||||||||||||
| SciCode | 53% | default | — | |||||||||||||||||||||||||||||
| ||||||||||||||||||||||||||||||||
| TerminalBench v2.1 | 81% | default | — | |||||||||||||||||||||||||||||
| ||||||||||||||||||||||||||||||||
| τ-bench Banking | 31% | default | — | |||||||||||||||||||||||||||||
| ||||||||||||||||||||||||||||||||
Pricing detail
cost beyond the standard rate · source: models.devLong context (>200K)
$2.00 / $9.00
input / output · per 1M
Pricing by serving provider
last-known · per 1M tokens · click a provider to expand| Serving provider | Input /1M | Output /1M | Endpoints | |||||||||||||
|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|
| Azure | $0.20 | $1.20 | 1 | |||||||||||||
| ||||||||||||||||
| OpenAI | $0.20 | $1.20 | 1 | |||||||||||||
| ||||||||||||||||
| OpenRouter | $0.20 | $1.20 | 1 | |||||||||||||
| ||||||||||||||||
Price history
input + output $/1M since we started tracking Input Output
Input down 80% since first tracked
Cost calculator
estimate your monthly spend on this model$136
estimated / month
Model IDs
copy the exact identifier for your platformazure/gpt-5.6-lunagpt-5.6-lunaopenai/gpt-5.6-luna
Frequently asked questions
GPT-5.6 Luna pricing, context and availabilityIs GPT-5.6 Luna still available?
GPT-5.6 Luna was retired on 2028-01-11. The pricing on this page is the last-known rate, kept for migration reference.
How much did GPT-5.6 Luna cost?
GPT-5.6 Luna's last-known pricing, before it was retired on 2028-01-11, was $0.20 per 1M input tokens and $1.20 per 1M output tokens.
What was the context window of GPT-5.6 Luna?
GPT-5.6 Luna had a 922K token context window and could return up to 128K output tokens.
What could GPT-5.6 Luna do?
GPT-5.6 Luna supported image input (vision), tool use, reasoning, prompt caching, structured output and web search.
Other OpenAI models
compare pricing across the OpenAI lineup Pricing verified against LiteLLM + OpenRouter + provider pages · refreshed daily
Every weekday
AI moves fast. Here's your debrief.
News, analysis, tools, and more.
For people who build with AI