Context
262K
Max output
262K
Serving providers
4
Cheapest input
$0.25 /1M
Prices are on-demand list rates. Enterprise commitments (committed-use discounts, provisioned throughput, negotiated rates) can be materially lower.

Benchmarks

independent evaluations · source: Artificial Analysis and Hugging Face leaderboards
Speed & latency
Output speed
131 t/s
tokens / second
Time to first token
2.30s
latency
Headline indices
45.7
of 100
21.3
of 100
Individual benchmarks · click a row for effort variants
BenchmarkBest scoreEffortCost / task
CritPt1%reasoning: false
EvalComputeProxy243.6default
GDPval987.7default
GPQA Diamond86%default
Humanity's Last Exam25%default
IFBench76%default
Long-Context Reasoning70%default
MMMU-Pro75%default
Omniscience-41.5default
Omniscience Accuracy0.2default
Omniscience Non Hallucination0.1default
SciCode42%default
SWE-bench Verified72%default
TerminalBench Hard31%default
TerminalBench v2.148%default
τ-bench Banking15%default
τ²-bench94%default

Pricing by serving provider

standard tier · per 1M tokens · click a provider to expand

Input pricing runs from $0.25 to $0.40 per 1M tokens across 4 providers, so the dearest route costs 60% more than the cheapest for the same model.

Serving providerInput /1MOutput /1MEndpoints
LibertaiCheapest$0.25$1.751
ComponentUnitStandardBatchCached
Text input/1M tok$0.2500
Text output/1M tok$1.75
OpenRouter$0.29$2.401
DeepInfra$0.29$2.401
Novita$0.40$3.201

Price history

input + output $/1M since we started tracking
Input Output
Input down 38% since first tracked
$0.000$0.500$1.00$1.50$2.00Mar 9Jun 22Aug 19Aug 29$1.75$0.250

Cost calculator

estimate your monthly spend on this model
$190
estimated / month

Model IDs

copy the exact identifier for your platform
libertai/qwen3.5-122b-a10bqwen/qwen3.5-122b-a10bdeepinfra/Qwen/Qwen3.5-122B-A10Bnovita/qwen/qwen3.5-122b-a10b

Frequently asked questions

Qwen3.5-122B-A10B pricing, context and availability

How much does Qwen3.5-122B-A10B cost?

Qwen3.5-122B-A10B costs $0.25 per 1M input tokens and $1.75 per 1M output tokens at its cheapest provider via Libertai. Across 4 serving providers, input prices range from $0.25 to $0.40 per 1M tokens.

What is the context window of Qwen3.5-122B-A10B?

Qwen3.5-122B-A10B accepts up to 262K tokens of context and can return up to 262K output tokens.

Which providers serve Qwen3.5-122B-A10B?

Qwen3.5-122B-A10B is available from 4 serving providers, each with its own pricing and model ID. Libertai is currently the cheapest.

What can Qwen3.5-122B-A10B do?

Qwen3.5-122B-A10B supports image input (vision), tool use, reasoning and structured output.

Other Alibaba models

compare pricing across the Alibaba lineup
Pricing verified against LiteLLM + OpenRouter + provider pages · refreshed daily

Every weekday

AI moves fast. Here's your debrief.

News, analysis, tools, and more.

For people who build with AI