Context
33K
Max output
n/a
Serving providers
1
Cheapest input
$0.05 /1M
Prices are on-demand list rates. Enterprise commitments (committed-use discounts, provisioned throughput, negotiated rates) can be materially lower.

Benchmarks

independent evaluations · source: Artificial Analysis
Speed & latency
Output speed
58 t/s
tokens / second
Time to first token
2.77s
latency
Headline indices
Individual benchmarks · click a row for effort variants
BenchmarkBest scoreEffortCost / task
CritPt0%default
GPQA Diamond47%default
Humanity's Last Exam4%default
IFBench24%default
LiveCodeBench28%default
Long-Context Reasoning0%default
Omniscience-66.7default
Omniscience Accuracy0.1default
Omniscience Non Hallucination0.1default
SciCode27%default
TerminalBench Hard5%default
τ²-bench32%default

Pricing by serving provider

standard tier · per 1M tokens · click a provider to expand
Serving providerInput /1MOutput /1MEndpoints
WandbCheapest$0.05$0.221
EndpointStd inStd outCached inBatch in / outModel ID
StandardCheapest$0.05$0.22wandb/OpenPipe/Qwen3-14B-Instruct

Price history

input + output $/1M since we started tracking
Input Output

Price history is accruing. We record a point each day a price changes; the full backfill lands shortly.

Cost calculator

estimate your monthly spend on this model
$28
estimated / month

Model IDs

copy the exact identifier for your platform
wandb/OpenPipe/Qwen3-14B-Instruct

Frequently asked questions

Qwen3 14B Instruct pricing, context and availability

How much does Qwen3 14B Instruct cost?

Qwen3 14B Instruct costs $0.05 per 1M input tokens and $0.22 per 1M output tokens at its cheapest provider via Wandb.

What is the context window of Qwen3 14B Instruct?

Qwen3 14B Instruct accepts up to 33K tokens of context.

Other Alibaba models

compare pricing across the Alibaba lineup
Pricing verified against LiteLLM + OpenRouter + provider pages · refreshed daily

Every weekday

AI moves fast. Here's your debrief.

News, analysis, tools, and more.

For people who build with AI