This model was retired on 2026-10-16. The pricing below is the last-known rate, kept for migration reference.

o4 MiniDeprecated

OpenAIReleasedApr 16, 2025KnowledgeMay 2024
VisionTool useReasoningPrompt cachingStructured outputWeb search
Context
200K
Max output
100K
Serving providers
5
Cheapest input
$1.00 /1M
Prices are on-demand list rates. Enterprise commitments (committed-use discounts, provisioned throughput, negotiated rates) can be materially lower.

Benchmarks

independent evaluations · source: Artificial Analysis
Speed & latency
Output speed
138 t/s
tokens / second
Time to first token
28.99s
latency
Headline indices
Intelligence Index
26.1
of 100
Individual benchmarks · click a row for effort variants
BenchmarkBest scoreEffortCost / task
CritPt1%default
GPQA Diamond78%default
Humanity's Last Exam17%default
IFBench69%default
LiveCodeBench86%default
Long-Context Reasoning60%default
MMLU-Pro83%default
MMMU-Pro69%default
Omniscience-35.7default
Omniscience Accuracy0.2default
Omniscience Non Hallucination0.2default
SciCode47%default
TerminalBench Hard15%default
τ²-bench56%default

Pricing by serving provider

last-known · per 1M tokens · click a provider to expand
Serving providerInput /1MOutput /1MEndpoints
Replicate$1.00$4.001
OpenAI$1.10$4.401
Vercel AI Gateway$1.10$4.401
Azure$1.10$4.401
OpenRouter$1.10$4.401

Price history

input + output $/1M since we started tracking
Input Output
Input down 9% since first tracked
$0.000$2.00$4.00$6.00Apr 16Jul 30Jan 12Jul 19$4.00$1.00

Cost calculator

estimate your monthly spend on this model
$520
estimated / month

Model IDs

copy the exact identifier for your platform
replicate/openai/o4-minio4-minivercel_ai_gateway/openai/o4-miniazure/o4-miniopenai/o4-mini

Frequently asked questions

o4 Mini pricing, context and availability

Is o4 Mini still available?

o4 Mini was retired on 2026-10-16. The pricing on this page is the last-known rate, kept for migration reference.

How much did o4 Mini cost?

o4 Mini's last-known pricing, before it was retired on 2026-10-16, was $1.00 per 1M input tokens and $4.00 per 1M output tokens.

What was the context window of o4 Mini?

o4 Mini had a 200K token context window and could return up to 100K output tokens.

What could o4 Mini do?

o4 Mini supported image input (vision), tool use, reasoning, prompt caching, structured output and web search.

Other OpenAI models

compare pricing across the OpenAI lineup

Popular comparisons

head-to-head pages featuring o4 Mini
Pricing verified against LiteLLM + OpenRouter + provider pages · refreshed daily

Every weekday

AI moves fast. Here's your debrief.

News, analysis, tools, and more.

For people who build with AI