Context
262K
Max output
n/a
Serving providers
1
Cheapest input
$0.10 /1M
Prices are on-demand list rates. Enterprise commitments (committed-use discounts, provisioned throughput, negotiated rates) can be materially lower.
Benchmarks
independent evaluations · source: Artificial AnalysisSpeed & latency
Output speed
43 t/s
tokens / second
Time to first token
9.23s
latency
Headline indices
Individual benchmarks · click a row for effort variants
| Benchmark | Best score | Effort | Cost / task | |||||||||||||
|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|
| CritPt | 1% | reasoning: false | — | |||||||||||||
| ||||||||||||||||
| EvalComputeProxy | 4852.7 | default | — | |||||||||||||
| ||||||||||||||||
| GDPval | 635.6 | default | — | |||||||||||||
| GPQA Diamond | 81% | default | — | |||||||||||||
| ||||||||||||||||
| Humanity's Last Exam | 13% | default | — | |||||||||||||
| ||||||||||||||||
| IFBench | 67% | default | — | |||||||||||||
| ||||||||||||||||
| Long-Context Reasoning | 59% | default | — | |||||||||||||
| ||||||||||||||||
| MMMU-Pro | 69% | default | — | |||||||||||||
| ||||||||||||||||
| Omniscience | -52.5 | default | — | |||||||||||||
| ||||||||||||||||
| SciCode | 28% | reasoning: false | — | |||||||||||||
| ||||||||||||||||
| TerminalBench Hard | 24% | default | — | |||||||||||||
| ||||||||||||||||
| TerminalBench v2.1 | 29% | default | — | |||||||||||||
| ||||||||||||||||
| τ-bench Banking | 8% | default | — | |||||||||||||
| ||||||||||||||||
| τ²-bench | 87% | default | — | |||||||||||||
| ||||||||||||||||
Pricing by serving provider
standard tier · per 1M tokens · click a provider to expand| Serving provider | Input /1M | Output /1M | Endpoints | |||||||||||||
|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|
| OpenRouterCheapest | $0.10 | $0.15 | 1 | |||||||||||||
| ||||||||||||||||
Price history
input + output $/1M since we started tracking Input Output
Price history is accruing. We record a point each day a price changes; the full backfill lands shortly.
Cost calculator
estimate your monthly spend on this model$32
estimated / month
Model IDs
copy the exact identifier for your platformqwen/qwen3.5-9b
Frequently asked questions
Qwen3.5-9B pricing, context and availabilityHow much does Qwen3.5-9B cost?
Qwen3.5-9B costs $0.10 per 1M input tokens and $0.15 per 1M output tokens at its cheapest provider via OpenRouter.
What is the context window of Qwen3.5-9B?
Qwen3.5-9B accepts up to 262K tokens of context.
Other Alibaba models
compare pricing across the Alibaba lineupQwen3 32B$0.08/1M · 8 providersQwQ 32B$0.15/1M · 7 providersQwen3 235B A22b Instruct 2507$0.09/1M · 6 providersQwen3 Next 80B A3B Thinking$0.10/1M · 6 providersQwen3 Next 80B A3B Instruct$0.10/1M · 6 providersQwen3 30B A3B$0.08/1M · 5 providersQwen3 235B A22B Thinking 2507$0.11/1M · 5 providersQwen3 235B A22B$0.18/1M · 5 providers
Pricing verified against LiteLLM + OpenRouter + provider pages · refreshed daily
Every weekday
AI moves fast. Here's your debrief.
News, analysis, tools, and more.
For people who build with AI