DeepSeek V4 Pro 0423 vs Qwen3.5 397B A17B

API pricing, context window and capabilities for DeepSeek V4 Pro 0423 and Qwen3.5 397B A17B, side by side. Prices are the cheapest serving route per 1M tokens, verified against LiteLLM and OpenRouter.

Verdict

computed from the published numbers

DeepSeek V4 Pro 0423 wins on coding, on intelligence and on context window.

Price
Tied
Coding
DeepSeek V4 Pro 0423
Intelligence
DeepSeek V4 Pro 0423
Context window
DeepSeek V4 Pro 0423

Key takeaways

who wins on what, from the published numbers

DeepSeek V4 Pro 0423 wins

  • Cheaper output ($0.87 vs $2.34 /1M)
  • Larger context window (1M vs 262K)
  • Higher max output (512K vs 66K)
  • More serving providers (10 vs 5)

Qwen3.5 397B A17B wins

  • Cheaper input ($0.43 vs $0.39 /1M)
  • Supports vision
Cheaper overall
DeepSeek V4 Pro 0423
$0.43 + $0.87 /1M
Larger context
DeepSeek V4 Pro 0423
1M
More providers
DeepSeek V4 Pro 0423
10 providers

Pricing & specs

cheapest serving route · per 1M tokens
DeepSeek V4 Pro 0423Qwen3.5 397B A17B
Input $/1M$0.43 $0.39 best
Output $/1M$0.87 best$2.34
Context window1M best262K
Max output512K best66K
Serving providers10 best5
Released2026-04-24 2026-02-16

Benchmarks

Artificial Analysis indices, side by side
DeepSeek V4 Pro 0423Qwen3.5 397B A17B
Intelligence53.2 best34.3
Coding68.8 best48.2
GPQA Diamond93% best89%
Output speed68 t/s 90 t/s best
Latency (TTFT)1.75s best2.27s

Cost calculator

choose a serving route for each side
Input tokens / month
Output tokens / month
DeepSeek V4 Pro 0423
via Tencent
$3.04
Qwen3.5 397B A17B
via OpenRouter
$4.29

Capabilities

feature support, side by side
FeatureDeepSeek V4 Pro 0423Qwen3.5 397B A17B
Vision
Tool use
Reasoning
Audio
Prompt caching
Structured output
Web search

Related comparisons

Frequently asked questions

Which is cheaper, DeepSeek V4 Pro 0423 or Qwen3.5 397B A17B?
Qwen3.5 397B A17B is cheaper for input ($0.43 vs $0.39 per 1M tokens). DeepSeek V4 Pro 0423 is cheaper for output ($0.87 vs $2.34 per 1M tokens).
What is the context window difference between DeepSeek V4 Pro 0423 and Qwen3.5 397B A17B?
DeepSeek V4 Pro 0423 has a 1M-token context window and Qwen3.5 397B A17B has 262K tokens, so DeepSeek V4 Pro 0423 can take a larger prompt.
How do DeepSeek V4 Pro 0423 and Qwen3.5 397B A17B compare on benchmarks?
On the tracked Artificial Analysis indices, DeepSeek V4 Pro 0423 leads on coding, and DeepSeek V4 Pro 0423 leads on intelligence.
Pricing verified against LiteLLM + OpenRouter + provider pages · refreshed daily

Every weekday

AI moves fast. Here's your debrief.

News, analysis, tools, and more.

For people who build with AI