Benchmarks
independent evaluations · source: Artificial Analysis| Benchmark | Best score | Effort | Cost / task | |
|---|---|---|---|---|
| CritPt | 0% | default | — | |
| EvalComputeProxy | 4.9 | default | — | |
| GPQA Diamond | 71% | default | — | |
| Humanity's Last Exam | 7% | default | — | |
| IFBench | 43% | default | — | |
| LiveCodeBench | 59% | default | — | |
| Long-Context Reasoning | 32% | default | — | |
| MMLU-Pro | 82% | default | — | |
| MMMU-Pro | 68% | default | — | |
| Omniscience | -52.7 | default | — | |
| Omniscience Accuracy | 0.2 | default | — | |
| Omniscience Non Hallucination | 0.1 | default | — | |
| SciCode | 36% | default | — | |
| TerminalBench Hard | 7% | default | — | |
| τ²-bench | 35% | default | — |
Pricing by serving provider
standard tier · per 1M tokens · click a provider to expandInput pricing runs from $0.20 to $0.40 per 1M tokens across 5 providers, so the dearest route costs 100% more than the cheapest for the same model.
| Serving provider | Input /1M | Output /1M | Endpoints | ||||||||||||||||
|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|
| DeepInfraCheapest | $0.20 | $0.88 | 1 | ||||||||||||||||
| |||||||||||||||||||
| OpenRouter | $0.21 | $1.90 | 1 | ||||||||||||||||
| |||||||||||||||||||
| Fireworks AI | $0.22 | $0.88 | 1 | ||||||||||||||||
| |||||||||||||||||||
| Novita | $0.30 | $1.50 | 1 | ||||||||||||||||
| |||||||||||||||||||
| Dashscope | $0.40 | $1.60 | 1 | ||||||||||||||||
| |||||||||||||||||||
Price history
input + output $/1M since we started trackingCost calculator
estimate your monthly spend on this modelModel IDs
copy the exact identifier for your platformFrequently asked questions
Qwen3 VL 235B A22B Instruct pricing, context and availabilityHow much does Qwen3 VL 235B A22B Instruct cost?
Qwen3 VL 235B A22B Instruct costs $0.20 per 1M input tokens and $0.88 per 1M output tokens at its cheapest provider via DeepInfra. Across 5 serving providers, input prices range from $0.20 to $0.40 per 1M tokens.
What is the context window of Qwen3 VL 235B A22B Instruct?
Qwen3 VL 235B A22B Instruct accepts up to 262K tokens of context and can return up to 262K output tokens.
Which providers serve Qwen3 VL 235B A22B Instruct?
Qwen3 VL 235B A22B Instruct is available from 5 serving providers, each with its own pricing and model ID. DeepInfra is currently the cheapest.
What can Qwen3 VL 235B A22B Instruct do?
Qwen3 VL 235B A22B Instruct supports image input (vision), tool use, prompt caching and structured output.
Other Alibaba models
compare pricing across the Alibaba lineupEvery weekday
AI moves fast. Here's your debrief.
News, analysis, tools, and more.
For people who build with AI