GLM 5.1 vs Qwen3.5 397B A17B
API pricing, context window and capabilities for GLM 5.1 and Qwen3.5 397B A17B, side by side. Prices are the cheapest serving route per 1M tokens, verified against LiteLLM and OpenRouter.
Verdict
computed from the published numbersGLM 5.1 wins on coding and on intelligence. Qwen3.5 397B A17B wins on price and on context window.
Price
Qwen3.5 397B A17B
Coding
GLM 5.1
Intelligence
GLM 5.1
Context window
Qwen3.5 397B A17B
Key takeaways
who wins on what, from the published numbersGLM 5.1 wins
- Higher max output (131K vs 66K)
Qwen3.5 397B A17B wins
- Cheaper input ($1.05 vs $0.39 /1M)
- Cheaper output ($3.50 vs $2.34 /1M)
- Larger context window (205K vs 262K)
- Supports vision
Cheaper overall
Qwen3.5 397B A17B
$0.39 + $2.34 /1M
Larger context
Qwen3.5 397B A17B
262K
More providers
Tied
Pricing & specs
cheapest serving route · per 1M tokens| GLM 5.1 | Qwen3.5 397B A17B | |
|---|---|---|
| Input $/1M | $1.05 | $0.39 best |
| Output $/1M | $3.50 | $2.34 best |
| Context window | 205K | 262K best |
| Max output | 131K best | 66K |
| Serving providers | 5 | 5 |
| Released | 2026-04-07 | 2026-02-16 |
Benchmarks
Artificial Analysis indices, side by side| GLM 5.1 | Qwen3.5 397B A17B | |
|---|---|---|
| Intelligence | 41.0 best | 34.3 |
| Coding | 55.8 best | 48.2 |
| GPQA Diamond | 87% | 89% best |
| Output speed | 56 t/s | 90 t/s best |
| Latency (TTFT) | 1.71s best | 2.27s |
Cost calculator
choose a serving route for each sideInput tokens / month
Output tokens / month
GLM 5.1
via DeepInfra
$8.75
Qwen3.5 397B A17B
via OpenRouter
$4.29
Capabilities
feature support, side by side| Feature | GLM 5.1 | Qwen3.5 397B A17B |
|---|---|---|
| Vision | — | ✓ |
| Tool use | ✓ | ✓ |
| Reasoning | ✓ | ✓ |
| Audio | — | — |
| Prompt caching | ✓ | ✓ |
| Structured output | ✓ | ✓ |
| Web search | — | — |
Related comparisons
Frequently asked questions
Which is cheaper, GLM 5.1 or Qwen3.5 397B A17B?
Qwen3.5 397B A17B is cheaper for input ($1.05 vs $0.39 per 1M tokens). Qwen3.5 397B A17B is cheaper for output ($3.50 vs $2.34 per 1M tokens).
What is the context window difference between GLM 5.1 and Qwen3.5 397B A17B?
GLM 5.1 has a 205K-token context window and Qwen3.5 397B A17B has 262K tokens, so Qwen3.5 397B A17B can take a larger prompt.
How do GLM 5.1 and Qwen3.5 397B A17B compare on benchmarks?
On the tracked Artificial Analysis indices, GLM 5.1 leads on coding, and GLM 5.1 leads on intelligence.
Pricing verified against LiteLLM + OpenRouter + provider pages · refreshed daily
Every weekday
AI moves fast. Here's your debrief.
News, analysis, tools, and more.
For people who build with AI