Context
32K
Max output
32K
Serving providers
1
Cheapest input
$0.55 /1M
Prices are on-demand list rates. Enterprise commitments (committed-use discounts, provisioned throughput, negotiated rates) can be materially lower.
Pricing by serving provider
standard tier · per 1M tokens · click a provider to expand| Serving provider | Input /1M | Output /1M | Endpoints | |||||||||||||
|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|
| NovitaCheapest | $0.55 | $1.66 | 1 | |||||||||||||
| ||||||||||||||||
Price history
input + output $/1M since we started tracking Input Output
Price history is accruing. We record a point each day a price changes; the full backfill lands shortly.
Cost calculator
estimate your monthly spend on this model$243
estimated / month
Model IDs
copy the exact identifier for your platformnovita/thudm/glm-4-32b-0414
Frequently asked questions
GLM 4 32B 0414 pricing, context and availabilityHow much does GLM 4 32B 0414 cost?
GLM 4 32B 0414 costs $0.55 per 1M input tokens and $1.66 per 1M output tokens at its cheapest provider via Novita.
What is the context window of GLM 4 32B 0414?
GLM 4 32B 0414 accepts up to 32K tokens of context and can return up to 32K output tokens.
What can GLM 4 32B 0414 do?
GLM 4 32B 0414 supports tool use and structured output.
Other Zhipu models
compare pricing across the Zhipu lineup Pricing verified against LiteLLM + OpenRouter + provider pages · refreshed daily
Every weekday
AI moves fast. Here's your debrief.
News, analysis, tools, and more.
For people who build with AI