Context
1M
Max output
1M
Serving providers
10
Cheapest input
$0.61 /1M
Prices are on-demand list rates. Enterprise commitments (committed-use discounts, provisioned throughput, negotiated rates) can be materially lower.

Benchmarks

independent evaluations · source: ARC Prize, Artificial Analysis and Hugging Face leaderboards
Speed & latency
Output speed
70 t/s
tokens / second
Time to first token
1.56s
latency
Headline indices
68.8
of 100
45.7
of 100
Individual benchmarks · click a row for effort variants
BenchmarkBest scoreEffortCost / task
AA-Briefcase38%
APEX Agents34%default
ARC-AGI-223%default$0.250
Automation Bench0.3default
CritPt21%default
EvalComputeProxy3533.8default
GDPval1505.2default
GPQA Diamond89%default
Harvey Lab0.9default
Humanity's Last Exam41%default
IFBench73%default
IT-Bench SRE43%default
Long-Context Reasoning77%default
Mlcr Overall0.1default
Omniscience4.4default
Omniscience Accuracy0.2default
Omniscience Non Hallucination0.7default
SciCode50%default
SWE-bench Pro62%default
TerminalBench Hard51%default
TerminalBench v2.178%default
τ-bench Banking35%default
τ²-bench99%default

Pricing by serving provider

standard tier · per 1M tokens · click a provider to expand

Input pricing runs from $0.61 to $1.40 per 1M tokens across 10 providers, so the dearest route costs 130% more than the cheapest for the same model.

Serving providerInput /1MOutput /1MEndpoints
Scx AICheapest$0.61$1.981
EndpointStd inStd outCached inBatch in / outModel ID
StandardCheapest$0.61$1.98$0.22scx-ai/GLM-5.2
DeepInfra$0.75$2.401
Wandb$0.76$2.421
OpenRouter$1.19$3.741
Databricks$1.40$4.401
Cloudflare$1.40$4.401
Dashscope$1.40$4.401
Mistral$1.40$4.401
Databricks is dearest at $1.40 / $4.40

Price history

input + output $/1M since we started tracking
Input Output
Input down 56% since first tracked
$0.000$2.00$4.00$6.00Jun 24Aug 7Aug 21Aug 28$1.98$0.610

Cost calculator

estimate your monthly spend on this model
$280
estimated / month

Model IDs

copy the exact identifier for your platform
scx-ai/GLM-5.2deepinfra/zai-org/GLM-5.2wandb/zai-org/GLM-5.2z-ai/glm-5.2databricks/databricks-glm-5-2cloudflare/@cf/zai-org/glm-5.2dashscope/glm-5.2mistral/glm-5-2together_ai/zai-org/GLM-5.2novita/zai-org/glm-5.2

Frequently asked questions

GLM 5.2 pricing, context and availability

How much does GLM 5.2 cost?

GLM 5.2 costs $0.61 per 1M input tokens and $1.98 per 1M output tokens at its cheapest provider via Scx AI. Across 10 serving providers, input prices range from $0.61 to $1.40 per 1M tokens.

What is the context window of GLM 5.2?

GLM 5.2 accepts up to 1M tokens of context and can return up to 1M output tokens.

Which providers serve GLM 5.2?

GLM 5.2 is available from 10 serving providers, each with its own pricing and model ID. Scx AI is currently the cheapest.

What can GLM 5.2 do?

GLM 5.2 supports tool use, reasoning, prompt caching and structured output.

Other Zhipu models

compare pricing across the Zhipu lineup

Popular comparisons

head-to-head pages featuring GLM 5.2
Pricing verified against LiteLLM + OpenRouter + provider pages · refreshed daily

Every weekday

AI moves fast. Here's your debrief.

News, analysis, tools, and more.

For people who build with AI