This model was retired on 2027-04-14. The pricing below is the last-known rate, kept for migration reference.

GPT-4.1Deprecated

OpenAI
VisionTool usePrompt cachingStructured outputWeb search
Context
1M
Max output
33K
Serving providers
5
Cheapest input
$2.00 /1M
Prices are on-demand list rates. Enterprise commitments (committed-use discounts, provisioned throughput, negotiated rates) can be materially lower.

Benchmarks

independent evaluations · source: Artificial Analysis
Speed & latency
Output speed
142 t/s
tokens / second
Time to first token
0.90s
latency
Headline indices
Intelligence Index
19.6
of 100
Individual benchmarks · click a row for effort variants
BenchmarkBest scoreEffortCost / task
CritPt0%default
GPQA Diamond67%default
Humanity's Last Exam4%default
IFBench43%default
LiveCodeBench46%default
Long-Context Reasoning64%default
MMLU-Pro81%default
MMMU-Pro61%default
Omniscience-39.6default
Omniscience Accuracy0.3default
Omniscience Non Hallucination0.1default
SciCode38%default
TerminalBench Hard14%default
τ²-bench47%default

Pricing by serving provider

last-known · per 1M tokens · click a provider to expand
Serving providerInput /1MOutput /1MEndpoints
Azure$2.00$8.001
OpenAI$2.00$8.001
Vercel AI Gateway$2.00$8.001
Replicate$2.00$8.001
OpenRouter$2.00$8.001

Price history

input + output $/1M since we started tracking
Input Output
$0.000$2.00$4.00$6.00$8.00Apr 14Jul 30Sep 1Jun 22$8.00$2.00

Cost calculator

estimate your monthly spend on this model
$1,040
estimated / month

Model IDs

copy the exact identifier for your platform
azure/gpt-4.1gpt-4.1vercel_ai_gateway/openai/gpt-4.1replicate/openai/gpt-4.1openai/gpt-4.1

Frequently asked questions

GPT-4.1 pricing, context and availability

Is GPT-4.1 still available?

GPT-4.1 was retired on 2027-04-14. The pricing on this page is the last-known rate, kept for migration reference.

How much did GPT-4.1 cost?

GPT-4.1's last-known pricing, before it was retired on 2027-04-14, was $2.00 per 1M input tokens and $8.00 per 1M output tokens.

What was the context window of GPT-4.1?

GPT-4.1 had a 1M token context window and could return up to 33K output tokens.

What could GPT-4.1 do?

GPT-4.1 supported image input (vision), tool use, prompt caching, structured output and web search.

Other OpenAI models

compare pricing across the OpenAI lineup

Popular comparisons

head-to-head pages featuring GPT-4.1
Pricing verified against LiteLLM + OpenRouter + provider pages · refreshed daily

Every weekday

AI moves fast. Here's your debrief.

News, analysis, tools, and more.

For people who build with AI