This model was retired on 2027-02-09. The pricing below is the last-known rate, kept for migration reference.
Context
410K
Max output
128K
Serving providers
6
Cheapest input
$1.25 /1M
Prices are on-demand list rates. Enterprise commitments (committed-use discounts, provisioned throughput, negotiated rates) can be materially lower.
Benchmarks
independent evaluations · source: Artificial AnalysisSpeed & latency
Output speed
94 t/s
tokens / second
Time to first token
101.26s
latency
Headline indices
Intelligence Index
35.3
of 100
Coding Index
37.8
of 100
Agentic Index
26.5
of 100
Individual benchmarks · click a row for effort variants
| Benchmark | Best score | Effort | Cost / task | |||||||||||||||||||||
|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|
| CritPt | 6% | default | — | |||||||||||||||||||||
| ||||||||||||||||||||||||
| GDPval | 1082.9 | default | — | |||||||||||||||||||||
| GPQA Diamond | 85% | default | — | |||||||||||||||||||||
| ||||||||||||||||||||||||
| Humanity's Last Exam | 28% | default | — | |||||||||||||||||||||
| ||||||||||||||||||||||||
| IFBench | 73% | default | — | |||||||||||||||||||||
| ||||||||||||||||||||||||
| LiveCodeBench | 85% | default | — | |||||||||||||||||||||
| ||||||||||||||||||||||||
| Long-Context Reasoning | 76% | medium | — | |||||||||||||||||||||
| ||||||||||||||||||||||||
| MMLU-Pro | 87% | default | — | |||||||||||||||||||||
| ||||||||||||||||||||||||
| MMMU-Pro | 74% | medium | — | |||||||||||||||||||||
| ||||||||||||||||||||||||
| Omniscience | -8.7 | default | — | |||||||||||||||||||||
| ||||||||||||||||||||||||
| Omniscience Accuracy | 0.4 | default | — | |||||||||||||||||||||
| ||||||||||||||||||||||||
| Omniscience Non Hallucination | 0.2 | low | — | |||||||||||||||||||||
| ||||||||||||||||||||||||
| SciCode | 43% | default | — | |||||||||||||||||||||
| ||||||||||||||||||||||||
| TerminalBench Hard | 38% | medium | — | |||||||||||||||||||||
| ||||||||||||||||||||||||
| TerminalBench v2.1 | 35% | default | — | |||||||||||||||||||||
| τ-bench Banking | 22% | default | — | |||||||||||||||||||||
| τ²-bench | 87% | medium | — | |||||||||||||||||||||
| ||||||||||||||||||||||||
Pricing by serving provider
last-known · per 1M tokens · click a provider to expand| Serving provider | Input /1M | Output /1M | Endpoints | |||||||||||||||||||||
|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|
| Databricks | $1.25 | $10.00 | 1 | |||||||||||||||||||||
| ||||||||||||||||||||||||
| Azure | $1.25 | $10.00 | 1 | |||||||||||||||||||||
| ||||||||||||||||||||||||
| OpenAI | $1.25 | $10.00 | 1 | |||||||||||||||||||||
| ||||||||||||||||||||||||
| Gmi | $1.25 | $10.00 | 1 | |||||||||||||||||||||
| ||||||||||||||||||||||||
| Replicate | $1.25 | $10.00 | 1 | |||||||||||||||||||||
| ||||||||||||||||||||||||
| OpenRouter | $1.25 | $10.00 | 1 | |||||||||||||||||||||
| ||||||||||||||||||||||||
Price history
input + output $/1M since we started tracking Input Output
Cost calculator
estimate your monthly spend on this model$1,050
estimated / month
Model IDs
copy the exact identifier for your platformdatabricks/databricks-gpt-5azure/gpt-5gpt-5gmi/openai/gpt-5replicate/openai/gpt-5openai/gpt-5
Frequently asked questions
GPT-5 pricing, context and availabilityIs GPT-5 still available?
GPT-5 was retired on 2027-02-09. The pricing on this page is the last-known rate, kept for migration reference.
How much did GPT-5 cost?
GPT-5's last-known pricing, before it was retired on 2027-02-09, was $1.25 per 1M input tokens and $10.00 per 1M output tokens.
What was the context window of GPT-5?
GPT-5 had a 410K token context window and could return up to 128K output tokens.
What could GPT-5 do?
GPT-5 supported image input (vision), tool use, reasoning, prompt caching, structured output and web search.
Other OpenAI models
compare pricing across the OpenAI lineupPopular comparisons
head-to-head pages featuring GPT-5 Pricing verified against LiteLLM + OpenRouter + provider pages · refreshed daily
Every weekday
AI moves fast. Here's your debrief.
News, analysis, tools, and more.
For people who build with AI