This model was retired on 2026-02-06. The pricing below is the last-known rate, kept for migration reference.

Qwen3 235B A22b FP8 TputDeprecated

Alibaba
Context
40K
Max output
n/a
Serving providers
1
Cheapest input
$0.20 /1M
Prices are on-demand list rates. Enterprise commitments (committed-use discounts, provisioned throughput, negotiated rates) can be materially lower.

Pricing by serving provider

last-known · per 1M tokens · click a provider to expand
Serving providerInput /1MOutput /1MEndpoints
Together AI$0.20$0.601

Price history

input + output $/1M since we started tracking
Input Output
$0.000$0.200$0.400$0.600Aug 14Jun 22$0.600$0.200

Cost calculator

estimate your monthly spend on this model
$88
estimated / month

Model IDs

copy the exact identifier for your platform
together_ai/Qwen/Qwen3-235B-A22B-fp8-tput

Frequently asked questions

Qwen3 235B A22b FP8 Tput pricing, context and availability

Is Qwen3 235B A22b FP8 Tput still available?

Qwen3 235B A22b FP8 Tput was retired on 2026-02-06. The pricing on this page is the last-known rate, kept for migration reference.

How much did Qwen3 235B A22b FP8 Tput cost?

Qwen3 235B A22b FP8 Tput's last-known pricing, before it was retired on 2026-02-06, was $0.20 per 1M input tokens and $0.60 per 1M output tokens.

What was the context window of Qwen3 235B A22b FP8 Tput?

Qwen3 235B A22b FP8 Tput had a 40K token context window.

Other Alibaba models

compare pricing across the Alibaba lineup
Pricing verified against LiteLLM + OpenRouter + provider pages · refreshed daily

Every weekday

AI moves fast. Here's your debrief.

News, analysis, tools, and more.

For people who build with AI