Context
1M
Max output
1M
Serving providers
5
Cheapest input
$1.25 /1M
Prices are on-demand list rates. Enterprise commitments (committed-use discounts, provisioned throughput, negotiated rates) can be materially lower.

Benchmarks

independent evaluations · source: Artificial Analysis
Speed & latency
Output speed
203 t/s
tokens / second
Time to first token
2.27s
latency
Headline indices
66.0
of 100
30.9
of 100
Individual benchmarks · click a row for effort variants
BenchmarkBest scoreEffortCost / task
Aa Analyst Agent0.2default
Automation Bench0.3default
CritPt13%default
GDPval1271default
GPQA Diamond92%default
Harvey Lab0.8default
Humanity's Last Exam41%default
IFBench81%default
IT-Bench SRE42%default
Long-Context Reasoning75%default
Omniscience13.5default
Omniscience Accuracy0.3default
Omniscience Non Hallucination0.7default
SciCode49%default
TerminalBench Hard51%default
TerminalBench v2.175%default
τ-bench Banking12%default
τ²-bench95%default

Pricing by serving provider

standard tier · per 1M tokens · click a provider to expand

Input pricing runs from $1.25 to $2.50 per 1M tokens across 5 providers, so the dearest route costs 100% more than the cheapest for the same model.

Serving providerInput /1MOutput /1MEndpoints
NovitaCheapest$1.25$3.751
EndpointStd inStd outCached inBatch in / outModel ID
StandardCheapest$1.25$3.75$0.25novita/qwen/qwen3.7-max
OpenRouter$1.48$4.421
Together AI$2.50$7.501
Dashscope$2.50$7.501
DeepInfra$2.50$7.501

Price history

input + output $/1M since we started tracking
Input Output
Input down 15% since first tracked
$0.000$2.00$4.00$6.00Jul 19Jul 31Aug 26Aug 28$3.75$1.25

Cost calculator

estimate your monthly spend on this model
$550
estimated / month

Model IDs

copy the exact identifier for your platform
novita/qwen/qwen3.7-maxqwen/qwen3.7-maxtogether_ai/Qwen/Qwen3.7-Maxdashscope/qwen3.7-maxdeepinfra/Qwen/Qwen3.7-Max

Frequently asked questions

Qwen3.7 Max pricing, context and availability

How much does Qwen3.7 Max cost?

Qwen3.7 Max costs $1.25 per 1M input tokens and $3.75 per 1M output tokens at its cheapest provider via Novita. Across 5 serving providers, input prices range from $1.25 to $2.50 per 1M tokens.

What is the context window of Qwen3.7 Max?

Qwen3.7 Max accepts up to 1M tokens of context and can return up to 1M output tokens.

Which providers serve Qwen3.7 Max?

Qwen3.7 Max is available from 5 serving providers, each with its own pricing and model ID. Novita is currently the cheapest.

What can Qwen3.7 Max do?

Qwen3.7 Max supports tool use, reasoning, prompt caching and structured output.

Other Alibaba models

compare pricing across the Alibaba lineup
Pricing verified against LiteLLM + OpenRouter + provider pages · refreshed daily

Every weekday

AI moves fast. Here's your debrief.

News, analysis, tools, and more.

For people who build with AI