Context
262K
Max output
262K
Serving providers
7
Cheapest input
$0.10 /1M
Prices are on-demand list rates. Enterprise commitments (committed-use discounts, provisioned throughput, negotiated rates) can be materially lower.

Benchmarks

independent evaluations · source: Artificial Analysis and Hugging Face leaderboards
Speed & latency
Output speed
121 t/s
tokens / second
Time to first token
2.05s
latency
Headline indices
41.9
of 100
21.6
of 100
Individual benchmarks · click a row for effort variants
BenchmarkBest scoreEffortCost / task
AIME 202693%default
CritPt0%default
EvalComputeProxy336.1reasoning: false
GDPval1055.3default
GPQA Diamond84%default
Humanity's Last Exam22%default
IFBench64%default
Long-Context Reasoning67%default
MMMU-Pro75%default
Omniscience-22.2default
Omniscience Accuracy0.2default
Omniscience Non Hallucination0.5default
SciCode36%default
SWE-bench Pro50%default
SWE-bench Verified73%default
TerminalBench Hard35%default
TerminalBench v2.145%default
τ-bench Banking9%default
τ²-bench95%default

Pricing by serving provider

standard tier · per 1M tokens · click a provider to expand

Input pricing runs from $0.10 to $0.25 per 1M tokens across 7 providers, so the dearest route costs 150% more than the cheapest for the same model.

Serving providerInput /1MOutput /1MEndpoints
DeepInfraCheapest$0.10$0.951
EndpointStd inStd outCached inBatch in / outModel ID
StandardCheapest$0.10$0.95deepinfra/Qwen/Qwen3.6-35B-A3B
OpenRouterCheapest$0.10$0.901
Pinstripes$0.14$0.451
Libertai$0.15$0.501
Novita$0.25$1.491
Wandb$0.25$1.251
Scaleway$0.25$1.501

Price history

input + output $/1M since we started tracking
Input Output
Input down 60% since first tracked
$0.000$0.500$1.00$1.50Jun 12Jun 22Aug 7Aug 28$0.450$0.100

Cost calculator

estimate your monthly spend on this model
$96
estimated / month

Model IDs

copy the exact identifier for your platform
deepinfra/Qwen/Qwen3.6-35B-A3Bqwen/qwen3.6-35b-a3bpinstripes/ps/qwen3.6-35b-a3blibertai/qwen3.6-35b-a3bnovita/qwen/qwen3.6-35b-a3bwandb/Qwen/Qwen3.6-35B-A3Bscaleway/qwen/qwen3.6-35b-a3b

Frequently asked questions

Qwen3.6 35B A3B pricing, context and availability

How much does Qwen3.6 35B A3B cost?

Qwen3.6 35B A3B costs $0.10 per 1M input tokens and $0.45 per 1M output tokens at its cheapest provider via DeepInfra. Across 7 serving providers, input prices range from $0.10 to $0.25 per 1M tokens.

What is the context window of Qwen3.6 35B A3B?

Qwen3.6 35B A3B accepts up to 262K tokens of context and can return up to 262K output tokens.

Which providers serve Qwen3.6 35B A3B?

Qwen3.6 35B A3B is available from 7 serving providers, each with its own pricing and model ID. DeepInfra is currently the cheapest.

What can Qwen3.6 35B A3B do?

Qwen3.6 35B A3B supports image input (vision), tool use, reasoning and structured output.

Other Alibaba models

compare pricing across the Alibaba lineup
Pricing verified against LiteLLM + OpenRouter + provider pages · refreshed daily

Every weekday

AI moves fast. Here's your debrief.

News, analysis, tools, and more.

For people who build with AI