Context
256K
Max output
n/a
Serving providers
1
Cheapest input
$0.12 /1M
Prices are on-demand list rates. Enterprise commitments (committed-use discounts, provisioned throughput, negotiated rates) can be materially lower.

Pricing by serving provider

standard tier · per 1M tokens · click a provider to expand
Serving providerInput /1MOutput /1MEndpoints
OpenRouterCheapest$0.12$1.361
EndpointStd inStd outCached inBatch in / outModel ID
StandardCheapest$0.12$1.36qwen/qwen3-vl-8b-thinking

Price history

input + output $/1M since we started tracking
Input Output

Price history is accruing. We record a point each day a price changes; the full backfill lands shortly.

Cost calculator

estimate your monthly spend on this model
$133
estimated / month

Model IDs

copy the exact identifier for your platform
qwen/qwen3-vl-8b-thinking

Frequently asked questions

Qwen3 VL 8B Thinking pricing, context and availability

How much does Qwen3 VL 8B Thinking cost?

Qwen3 VL 8B Thinking costs $0.12 per 1M input tokens and $1.36 per 1M output tokens at its cheapest provider via OpenRouter.

What is the context window of Qwen3 VL 8B Thinking?

Qwen3 VL 8B Thinking accepts up to 256K tokens of context.

Other Alibaba models

compare pricing across the Alibaba lineup
Pricing verified against LiteLLM + OpenRouter + provider pages · refreshed daily

Every weekday

AI moves fast. Here's your debrief.

News, analysis, tools, and more.

For people who build with AI