Context
131K
Max output
n/a
Serving providers
1
Cheapest input
$0.15 /1M
Prices are on-demand list rates. Enterprise commitments (committed-use discounts, provisioned throughput, negotiated rates) can be materially lower.
Pricing by serving provider
standard tier · per 1M tokens · click a provider to expand| Serving provider | Input /1M | Output /1M | Endpoints | |||||||||||||
|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|
| DeepInfraCheapest | $0.15 | $0.60 | 1 | |||||||||||||
| ||||||||||||||||
Price history
input + output $/1M since we started tracking Input Output
Price history is accruing. We record a point each day a price changes; the full backfill lands shortly.
Cost calculator
estimate your monthly spend on this model$78
estimated / month
Model IDs
copy the exact identifier for your platformdeepinfra/openai/gpt-oss-120b-Turbo
Frequently asked questions
GPT OSS 120B Turbo pricing, context and availabilityHow much does GPT OSS 120B Turbo cost?
GPT OSS 120B Turbo costs $0.15 per 1M input tokens and $0.60 per 1M output tokens at its cheapest provider via DeepInfra.
What is the context window of GPT OSS 120B Turbo?
GPT OSS 120B Turbo accepts up to 131K tokens of context.
What can GPT OSS 120B Turbo do?
GPT OSS 120B Turbo supports tool use, reasoning and structured output.
Other OpenAI models
compare pricing across the OpenAI lineup Pricing verified against LiteLLM + OpenRouter + provider pages · refreshed daily
Every weekday
AI moves fast. Here's your debrief.
News, analysis, tools, and more.
For people who build with AI