Context
32K
Max output
8K
Serving providers
2
Cheapest input
$0.36 /1M
Prices are on-demand list rates. Enterprise commitments (committed-use discounts, provisioned throughput, negotiated rates) can be materially lower.

Pricing by serving provider

standard tier · per 1M tokens · click a provider to expand

Input pricing runs from $0.36 to $0.38 per 1M tokens across 2 providers, so the dearest route costs 6% more than the cheapest for the same model.

Serving providerInput /1MOutput /1MEndpoints
OpenRouterCheapest$0.36$0.401
EndpointStd inStd outCached inBatch in / outModel ID
StandardCheapest$0.36$0.40qwen/qwen-2.5-72b-instruct
Novita$0.38$0.401

Price history

input + output $/1M since we started tracking
Input Output
$0.000$0.100$0.200$0.300$0.400Jul 30Jan 13Jun 22Jul 19$0.400$0.360

Cost calculator

estimate your monthly spend on this model
$104
estimated / month

Model IDs

copy the exact identifier for your platform
qwen/qwen-2.5-72b-instructnovita/qwen/qwen-2.5-72b-instruct

Frequently asked questions

Qwen2.5 72B Instruct pricing, context and availability

How much does Qwen2.5 72B Instruct cost?

Qwen2.5 72B Instruct costs $0.36 per 1M input tokens and $0.40 per 1M output tokens at its cheapest provider via OpenRouter. Across 2 serving providers, input prices range from $0.36 to $0.38 per 1M tokens.

What is the context window of Qwen2.5 72B Instruct?

Qwen2.5 72B Instruct accepts up to 32K tokens of context and can return up to 8K output tokens.

Which providers serve Qwen2.5 72B Instruct?

Qwen2.5 72B Instruct is available from 2 serving providers, each with its own pricing and model ID. OpenRouter is currently the cheapest.

What can Qwen2.5 72B Instruct do?

Qwen2.5 72B Instruct supports tool use and structured output.

Other Alibaba models

compare pricing across the Alibaba lineup
Pricing verified against LiteLLM + OpenRouter + provider pages · refreshed daily

Every weekday

AI moves fast. Here's your debrief.

News, analysis, tools, and more.

For people who build with AI