This model was retired on 2026-04-02. The pricing below is the last-known rate, kept for migration reference.
Context
128K
Max output
n/a
Serving providers
1
Cheapest input
$0.20 /1M
Prices are on-demand list rates. Enterprise commitments (committed-use discounts, provisioned throughput, negotiated rates) can be materially lower.
Pricing by serving provider
last-known · per 1M tokens · click a provider to expand| Serving provider | Input /1M | Output /1M | Endpoints | ||||||||||||||||
|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|
| Together AI | $0.20 | $1.10 | 1 | ||||||||||||||||
| |||||||||||||||||||
Price history
input + output $/1M since we started tracking Input Output
Cost calculator
estimate your monthly spend on this model$128
estimated / month
Model IDs
copy the exact identifier for your platformtogether_ai/zai-org/GLM-4.5-Air-FP8
Frequently asked questions
GLM 4.5 Air FP8 pricing, context and availabilityIs GLM 4.5 Air FP8 still available?
GLM 4.5 Air FP8 was retired on 2026-04-02. The pricing on this page is the last-known rate, kept for migration reference.
How much did GLM 4.5 Air FP8 cost?
GLM 4.5 Air FP8's last-known pricing, before it was retired on 2026-04-02, was $0.20 per 1M input tokens and $1.10 per 1M output tokens.
What was the context window of GLM 4.5 Air FP8?
GLM 4.5 Air FP8 had a 128K token context window.
What could GLM 4.5 Air FP8 do?
GLM 4.5 Air FP8 supported tool use and structured output.
Other Zhipu models
compare pricing across the Zhipu lineup Pricing verified against LiteLLM + OpenRouter + provider pages · refreshed daily
Every weekday
AI moves fast. Here's your debrief.
News, analysis, tools, and more.
For people who build with AI