This model was retired on 2026-10-23. The pricing below is the last-known rate, kept for migration reference.

GPT-3.5 Turbo 16kDeprecated

OpenAI
Tool usePrompt caching
Context
16K
Max output
4K
Serving providers
2
Cheapest input
$3.00 /1M
Prices are on-demand list rates. Enterprise commitments (committed-use discounts, provisioned throughput, negotiated rates) can be materially lower.

Pricing by serving provider

last-known · per 1M tokens · click a provider to expand
Serving providerInput /1MOutput /1MEndpoints
OpenAI$3.00$4.001
OpenRouter$3.00$4.001

Price history

input + output $/1M since we started tracking
Input Output
$0.000$1.00$2.00$3.00$4.00Sep 20Nov 23Jun 22$4.00$3.00

Cost calculator

estimate your monthly spend on this model
$920
estimated / month

Model IDs

copy the exact identifier for your platform
gpt-3.5-turbo-16kopenai/gpt-3.5-turbo-16k

Frequently asked questions

GPT-3.5 Turbo 16k pricing, context and availability

Is GPT-3.5 Turbo 16k still available?

GPT-3.5 Turbo 16k was retired on 2026-10-23. The pricing on this page is the last-known rate, kept for migration reference.

How much did GPT-3.5 Turbo 16k cost?

GPT-3.5 Turbo 16k's last-known pricing, before it was retired on 2026-10-23, was $3.00 per 1M input tokens and $4.00 per 1M output tokens.

What was the context window of GPT-3.5 Turbo 16k?

GPT-3.5 Turbo 16k had a 16K token context window and could return up to 4K output tokens.

What could GPT-3.5 Turbo 16k do?

GPT-3.5 Turbo 16k supported tool use and prompt caching.

Other OpenAI models

compare pricing across the OpenAI lineup
Pricing verified against LiteLLM + OpenRouter + provider pages · refreshed daily

Every weekday

AI moves fast. Here's your debrief.

News, analysis, tools, and more.

For people who build with AI