This model was retired on 2026-10-23. The pricing below is the last-known rate, kept for migration reference.
Context
16K
Max output
4K
Serving providers
2
Cheapest input
$3.00 /1M
Prices are on-demand list rates. Enterprise commitments (committed-use discounts, provisioned throughput, negotiated rates) can be materially lower.
Pricing by serving provider
last-known · per 1M tokens · click a provider to expand| Serving provider | Input /1M | Output /1M | Endpoints | ||||||||||||||||
|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|
| OpenAI | $3.00 | $4.00 | 1 | ||||||||||||||||
| |||||||||||||||||||
| OpenRouter | $3.00 | $4.00 | 1 | ||||||||||||||||
| |||||||||||||||||||
Price history
input + output $/1M since we started tracking Input Output
Cost calculator
estimate your monthly spend on this model$920
estimated / month
Model IDs
copy the exact identifier for your platformgpt-3.5-turbo-16kopenai/gpt-3.5-turbo-16k
Frequently asked questions
GPT-3.5 Turbo 16k pricing, context and availabilityIs GPT-3.5 Turbo 16k still available?
GPT-3.5 Turbo 16k was retired on 2026-10-23. The pricing on this page is the last-known rate, kept for migration reference.
How much did GPT-3.5 Turbo 16k cost?
GPT-3.5 Turbo 16k's last-known pricing, before it was retired on 2026-10-23, was $3.00 per 1M input tokens and $4.00 per 1M output tokens.
What was the context window of GPT-3.5 Turbo 16k?
GPT-3.5 Turbo 16k had a 16K token context window and could return up to 4K output tokens.
What could GPT-3.5 Turbo 16k do?
GPT-3.5 Turbo 16k supported tool use and prompt caching.
Other OpenAI models
compare pricing across the OpenAI lineup Pricing verified against LiteLLM + OpenRouter + provider pages · refreshed daily
Every weekday
AI moves fast. Here's your debrief.
News, analysis, tools, and more.
For people who build with AI