This model was retired on 2026-10-16. The pricing below is the last-known rate, kept for migration reference.
Context
200K
Max output
100K
Serving providers
5
Cheapest input
$1.00 /1M
Prices are on-demand list rates. Enterprise commitments (committed-use discounts, provisioned throughput, negotiated rates) can be materially lower.
Benchmarks
independent evaluations · source: Artificial AnalysisSpeed & latency
Output speed
138 t/s
tokens / second
Time to first token
28.99s
latency
Headline indices
Intelligence Index
26.1
of 100
Individual benchmarks · click a row for effort variants
| Benchmark | Best score | Effort | Cost / task | |
|---|---|---|---|---|
| CritPt | 1% | default | — | |
| GPQA Diamond | 78% | default | — | |
| Humanity's Last Exam | 17% | default | — | |
| IFBench | 69% | default | — | |
| LiveCodeBench | 86% | default | — | |
| Long-Context Reasoning | 60% | default | — | |
| MMLU-Pro | 83% | default | — | |
| MMMU-Pro | 69% | default | — | |
| Omniscience | -35.7 | default | — | |
| Omniscience Accuracy | 0.2 | default | — | |
| Omniscience Non Hallucination | 0.2 | default | — | |
| SciCode | 47% | default | — | |
| TerminalBench Hard | 15% | default | — | |
| τ²-bench | 56% | default | — |
Pricing by serving provider
last-known · per 1M tokens · click a provider to expand| Serving provider | Input /1M | Output /1M | Endpoints | ||||||||||||||||||||||||||
|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|
| Replicate | $1.00 | $4.00 | 1 | ||||||||||||||||||||||||||
| |||||||||||||||||||||||||||||
| OpenAI | $1.10 | $4.40 | 1 | ||||||||||||||||||||||||||
| |||||||||||||||||||||||||||||
| Vercel AI Gateway | $1.10 | $4.40 | 1 | ||||||||||||||||||||||||||
| |||||||||||||||||||||||||||||
| Azure | $1.10 | $4.40 | 1 | ||||||||||||||||||||||||||
| |||||||||||||||||||||||||||||
| OpenRouter | $1.10 | $4.40 | 1 | ||||||||||||||||||||||||||
| |||||||||||||||||||||||||||||
Price history
input + output $/1M since we started tracking Input Output
Input down 9% since first tracked
Cost calculator
estimate your monthly spend on this model$520
estimated / month
Model IDs
copy the exact identifier for your platformreplicate/openai/o4-minio4-minivercel_ai_gateway/openai/o4-miniazure/o4-miniopenai/o4-mini
Frequently asked questions
o4 Mini pricing, context and availabilityIs o4 Mini still available?
o4 Mini was retired on 2026-10-16. The pricing on this page is the last-known rate, kept for migration reference.
How much did o4 Mini cost?
o4 Mini's last-known pricing, before it was retired on 2026-10-16, was $1.00 per 1M input tokens and $4.00 per 1M output tokens.
What was the context window of o4 Mini?
o4 Mini had a 200K token context window and could return up to 100K output tokens.
What could o4 Mini do?
o4 Mini supported image input (vision), tool use, reasoning, prompt caching, structured output and web search.
Other OpenAI models
compare pricing across the OpenAI lineupPopular comparisons
head-to-head pages featuring o4 Mini Pricing verified against LiteLLM + OpenRouter + provider pages · refreshed daily
Every weekday
AI moves fast. Here's your debrief.
News, analysis, tools, and more.
For people who build with AI