Context
203K
Max output
131K
Serving providers
5
Cheapest input
$0.60 /1M
Prices are on-demand list rates. Enterprise commitments (committed-use discounts, provisioned throughput, negotiated rates) can be materially lower.

Benchmarks

independent evaluations · source: ARC Prize, Artificial Analysis and Hugging Face leaderboards
Speed & latency
Output speed
62 t/s
tokens / second
Time to first token
1.42s
latency
Headline indices
Individual benchmarks · click a row for effort variants
BenchmarkBest scoreEffortCost / task
AIME 202696%default
APEX Agents14%default
ARC-AGI-25%default$0.270
CritPt2%default
EvalComputeProxy621.5default
GPQA Diamond82%default
Humanity's Last Exam29%default
IFBench72%default
Long-Context Reasoning71%default
Omniscience0.3default
Omniscience Accuracy0.3default
Omniscience Non Hallucination0.6default
SciCode46%default
SWE-bench Verified78%default
TerminalBench Hard43%default
τ²-bench98%default

Pricing by serving provider

standard tier · per 1M tokens · click a provider to expand

Input pricing runs from $0.60 to $1.00 per 1M tokens across 5 providers, so the dearest route costs 67% more than the cheapest for the same model.

Serving providerInput /1MOutput /1MEndpoints
DeepInfraCheapest$0.60$2.081
EndpointStd inStd outCached inBatch in / outModel ID
StandardCheapest$0.60$2.08$0.12deepinfra/zai-org/GLM-5
OpenRouterCheapest$0.60$1.921
Baseten$0.95$3.151
Zai$1.00$3.201
Novita$1.00$3.201

Price history

input + output $/1M since we started tracking
Input Output
Input down 25% since first tracked
$0.000$1.00$2.00$3.00$4.00Mar 3Jun 22Jul 20Aug 28$1.92$0.600

Cost calculator

estimate your monthly spend on this model
$286
estimated / month

Model IDs

copy the exact identifier for your platform
deepinfra/zai-org/GLM-5z-ai/glm-5baseten/zai-org/GLM-5zai/glm-5novita/zai-org/glm-5

Frequently asked questions

GLM 5 pricing, context and availability

How much does GLM 5 cost?

GLM 5 costs $0.60 per 1M input tokens and $1.92 per 1M output tokens at its cheapest provider via DeepInfra. Across 5 serving providers, input prices range from $0.60 to $1.00 per 1M tokens.

What is the context window of GLM 5?

GLM 5 accepts up to 203K tokens of context and can return up to 131K output tokens.

Which providers serve GLM 5?

GLM 5 is available from 5 serving providers, each with its own pricing and model ID. DeepInfra is currently the cheapest.

What can GLM 5 do?

GLM 5 supports tool use, reasoning, prompt caching and structured output.

Other Zhipu models

compare pricing across the Zhipu lineup
Pricing verified against LiteLLM + OpenRouter + provider pages · refreshed daily

Every weekday

AI moves fast. Here's your debrief.

News, analysis, tools, and more.

For people who build with AI