Context
32K
Max output
32K
Serving providers
1
Cheapest input
$0.55 /1M
Prices are on-demand list rates. Enterprise commitments (committed-use discounts, provisioned throughput, negotiated rates) can be materially lower.

Pricing by serving provider

standard tier · per 1M tokens · click a provider to expand
Serving providerInput /1MOutput /1MEndpoints
NovitaCheapest$0.55$1.661
EndpointStd inStd outCached inBatch in / outModel ID
StandardCheapest$0.55$1.66novita/thudm/glm-4-32b-0414

Price history

input + output $/1M since we started tracking
Input Output

Price history is accruing. We record a point each day a price changes; the full backfill lands shortly.

Cost calculator

estimate your monthly spend on this model
$243
estimated / month

Model IDs

copy the exact identifier for your platform
novita/thudm/glm-4-32b-0414

Frequently asked questions

GLM 4 32B 0414 pricing, context and availability

How much does GLM 4 32B 0414 cost?

GLM 4 32B 0414 costs $0.55 per 1M input tokens and $1.66 per 1M output tokens at its cheapest provider via Novita.

What is the context window of GLM 4 32B 0414?

GLM 4 32B 0414 accepts up to 32K tokens of context and can return up to 32K output tokens.

What can GLM 4 32B 0414 do?

GLM 4 32B 0414 supports tool use and structured output.

Other Zhipu models

compare pricing across the Zhipu lineup
Pricing verified against LiteLLM + OpenRouter + provider pages · refreshed daily

Every weekday

AI moves fast. Here's your debrief.

News, analysis, tools, and more.

For people who build with AI