Context
262K
Max output
n/a
Serving providers
1
Cheapest input
$0.10 /1M
Prices are on-demand list rates. Enterprise commitments (committed-use discounts, provisioned throughput, negotiated rates) can be materially lower.

Benchmarks

independent evaluations · source: Artificial Analysis
Speed & latency
Output speed
43 t/s
tokens / second
Time to first token
9.23s
latency
Headline indices
28.7
of 100
7.4
of 100
Individual benchmarks · click a row for effort variants
BenchmarkBest scoreEffortCost / task
CritPt1%reasoning: false
EvalComputeProxy4852.7default
GDPval635.6default
GPQA Diamond81%default
Humanity's Last Exam13%default
IFBench67%default
Long-Context Reasoning59%default
MMMU-Pro69%default
Omniscience-52.5default
SciCode28%reasoning: false
TerminalBench Hard24%default
TerminalBench v2.129%default
τ-bench Banking8%default
τ²-bench87%default

Pricing by serving provider

standard tier · per 1M tokens · click a provider to expand
Serving providerInput /1MOutput /1MEndpoints
OpenRouterCheapest$0.10$0.151
EndpointStd inStd outCached inBatch in / outModel ID
StandardCheapest$0.10$0.15qwen/qwen3.5-9b

Price history

input + output $/1M since we started tracking
Input Output

Price history is accruing. We record a point each day a price changes; the full backfill lands shortly.

Cost calculator

estimate your monthly spend on this model
$32
estimated / month

Model IDs

copy the exact identifier for your platform
qwen/qwen3.5-9b

Frequently asked questions

Qwen3.5-9B pricing, context and availability

How much does Qwen3.5-9B cost?

Qwen3.5-9B costs $0.10 per 1M input tokens and $0.15 per 1M output tokens at its cheapest provider via OpenRouter.

What is the context window of Qwen3.5-9B?

Qwen3.5-9B accepts up to 262K tokens of context.

Other Alibaba models

compare pricing across the Alibaba lineup
Pricing verified against LiteLLM + OpenRouter + provider pages · refreshed daily

Every weekday

AI moves fast. Here's your debrief.

News, analysis, tools, and more.

For people who build with AI