Context
262K
Max output
262K
Serving providers
5
Cheapest input
$0.20 /1M
Prices are on-demand list rates. Enterprise commitments (committed-use discounts, provisioned throughput, negotiated rates) can be materially lower.

Benchmarks

independent evaluations · source: Artificial Analysis
Speed & latency
Output speed
47 t/s
tokens / second
Time to first token
2.71s
latency
Headline indices
Individual benchmarks · click a row for effort variants
BenchmarkBest scoreEffortCost / task
CritPt0%default
EvalComputeProxy4.9default
GPQA Diamond71%default
Humanity's Last Exam7%default
IFBench43%default
LiveCodeBench59%default
Long-Context Reasoning32%default
MMLU-Pro82%default
MMMU-Pro68%default
Omniscience-52.7default
Omniscience Accuracy0.2default
Omniscience Non Hallucination0.1default
SciCode36%default
TerminalBench Hard7%default
τ²-bench35%default

Pricing by serving provider

standard tier · per 1M tokens · click a provider to expand

Input pricing runs from $0.20 to $0.40 per 1M tokens across 5 providers, so the dearest route costs 100% more than the cheapest for the same model.

Serving providerInput /1MOutput /1MEndpoints
DeepInfraCheapest$0.20$0.881
EndpointStd inStd outCached inBatch in / outModel ID
StandardCheapest$0.20$0.88$0.11deepinfra/Qwen/Qwen3-VL-235B-A22B-Instruct
OpenRouter$0.21$1.901
Fireworks AI$0.22$0.881
Novita$0.30$1.501
Dashscope$0.40$1.601

Price history

input + output $/1M since we started tracking
Input Output
Input down 9% since first tracked
$0.000$0.500$1.00Dec 9Mar 3Aug 12Aug 28$0.880$0.200

Cost calculator

estimate your monthly spend on this model
$110
estimated / month

Model IDs

copy the exact identifier for your platform
deepinfra/Qwen/Qwen3-VL-235B-A22B-Instructqwen/qwen3-vl-235b-a22b-instructfireworks_ai/accounts/fireworks/models/qwen3-vl-235b-a22b-instructnovita/qwen/qwen3-vl-235b-a22b-instructdashscope/qwen3-vl-235b-a22b-instruct

Frequently asked questions

Qwen3 VL 235B A22B Instruct pricing, context and availability

How much does Qwen3 VL 235B A22B Instruct cost?

Qwen3 VL 235B A22B Instruct costs $0.20 per 1M input tokens and $0.88 per 1M output tokens at its cheapest provider via DeepInfra. Across 5 serving providers, input prices range from $0.20 to $0.40 per 1M tokens.

What is the context window of Qwen3 VL 235B A22B Instruct?

Qwen3 VL 235B A22B Instruct accepts up to 262K tokens of context and can return up to 262K output tokens.

Which providers serve Qwen3 VL 235B A22B Instruct?

Qwen3 VL 235B A22B Instruct is available from 5 serving providers, each with its own pricing and model ID. DeepInfra is currently the cheapest.

What can Qwen3 VL 235B A22B Instruct do?

Qwen3 VL 235B A22B Instruct supports image input (vision), tool use, prompt caching and structured output.

Other Alibaba models

compare pricing across the Alibaba lineup
Pricing verified against LiteLLM + OpenRouter + provider pages · refreshed daily

Every weekday

AI moves fast. Here's your debrief.

News, analysis, tools, and more.

For people who build with AI