Context
262K
Max output
66K
Serving providers
3
Cheapest input
$0.20 /1M
Prices are on-demand list rates. Enterprise commitments (committed-use discounts, provisioned throughput, negotiated rates) can be materially lower.

Benchmarks

independent evaluations · source: Artificial Analysis and Hugging Face leaderboards
Speed & latency
Output speed
76 t/s
tokens / second
Time to first token
5.61s
latency
Headline indices
Individual benchmarks · click a row for effort variants
BenchmarkBest scoreEffortCost / task
AIME 202691%default
CritPt1%default
EvalComputeProxy774.8default
GPQA Diamond86%default
Humanity's Last Exam24%default
IFBench76%default
IT-Bench SRE35%default
Long-Context Reasoning72%default
MMMU-Pro75%default
Omniscience-44.0default
Omniscience Accuracy0.2default
Omniscience Non Hallucination0.2reasoning: false
SciCode39%default
SWE-bench Verified72%default
TerminalBench Hard33%default
τ²-bench94%default

Pricing by serving provider

standard tier · per 1M tokens · click a provider to expand

Input pricing runs from $0.20 to $0.30 per 1M tokens across 3 providers, so the dearest route costs 54% more than the cheapest for the same model.

Serving providerInput /1MOutput /1MEndpoints
OpenRouterCheapest$0.20$1.561
EndpointStd inStd outCached inBatch in / outModel ID
StandardCheapest$0.20$1.56qwen/qwen3.5-27b
DeepInfra$0.26$2.601
Novita$0.30$2.401

Price history

input + output $/1M since we started tracking
Input Output
Input down 35% since first tracked
$0.000$1.00$2.00$3.00Mar 9Jun 22Jul 19Jul 24Aug 28$1.56$0.195

Cost calculator

estimate your monthly spend on this model
$164
estimated / month

Model IDs

copy the exact identifier for your platform
qwen/qwen3.5-27bdeepinfra/Qwen/Qwen3.5-27Bnovita/qwen/qwen3.5-27b

Frequently asked questions

Qwen3.5-27B pricing, context and availability

How much does Qwen3.5-27B cost?

Qwen3.5-27B costs $0.20 per 1M input tokens and $1.56 per 1M output tokens at its cheapest provider via OpenRouter. Across 3 serving providers, input prices range from $0.20 to $0.30 per 1M tokens.

What is the context window of Qwen3.5-27B?

Qwen3.5-27B accepts up to 262K tokens of context and can return up to 66K output tokens.

Which providers serve Qwen3.5-27B?

Qwen3.5-27B is available from 3 serving providers, each with its own pricing and model ID. OpenRouter is currently the cheapest.

What can Qwen3.5-27B do?

Qwen3.5-27B supports image input (vision), tool use, reasoning and structured output.

Other Alibaba models

compare pricing across the Alibaba lineup
Pricing verified against LiteLLM + OpenRouter + provider pages · refreshed daily

Every weekday

AI moves fast. Here's your debrief.

News, analysis, tools, and more.

For people who build with AI