Context
205K
Max output
131K
Serving providers
5
Cheapest input
$1.05 /1M
Prices are on-demand list rates. Enterprise commitments (committed-use discounts, provisioned throughput, negotiated rates) can be materially lower.

Benchmarks

independent evaluations · source: Artificial Analysis
Speed & latency
Output speed
56 t/s
tokens / second
Time to first token
1.71s
latency
Headline indices
55.8
of 100
30.6
of 100
Individual benchmarks · click a row for effort variants
BenchmarkBest scoreEffortCost / task
AA-Briefcase24%
CritPt5%default
EvalComputeProxy1461.7default
GDPval1257.8default
GPQA Diamond87%default
Humanity's Last Exam30%default
IFBench76%default
IT-Bench SRE40%default
Long-Context Reasoning68%default
Omniscience0.8default
Omniscience Accuracy0.3reasoning: false
Omniscience Non Hallucination0.7default
SciCode44%default
TerminalBench Hard43%default
TerminalBench v2.162%default
τ-bench Banking14%default
τ²-bench98%default

Pricing by serving provider

standard tier · per 1M tokens · click a provider to expand

Input pricing runs from $1.05 to $1.40 per 1M tokens across 5 providers, so the dearest route costs 33% more than the cheapest for the same model.

Serving providerInput /1MOutput /1MEndpoints
DeepInfraCheapest$1.05$3.501
EndpointStd inStd outCached inBatch in / outModel ID
StandardCheapest$1.05$3.50$0.20deepinfra/zai-org/GLM-5.1
OpenRouter$1.26$3.961
Novita$1.38$4.401
Zai$1.40$4.401
Dashscope$1.40$4.401

Price history

input + output $/1M since we started tracking
Input Output
$0.000$2.00$4.00$6.00Jun 27Aug 6Aug 15Aug 28$3.50$1.05

Cost calculator

estimate your monthly spend on this model
$490
estimated / month

Model IDs

copy the exact identifier for your platform
deepinfra/zai-org/GLM-5.1z-ai/glm-5.1novita/zai-org/glm-5.1zai/glm-5.1dashscope/glm-5.1

Frequently asked questions

GLM 5.1 pricing, context and availability

How much does GLM 5.1 cost?

GLM 5.1 costs $1.05 per 1M input tokens and $3.50 per 1M output tokens at its cheapest provider via DeepInfra. Across 5 serving providers, input prices range from $1.05 to $1.40 per 1M tokens.

What is the context window of GLM 5.1?

GLM 5.1 accepts up to 205K tokens of context and can return up to 131K output tokens.

Which providers serve GLM 5.1?

GLM 5.1 is available from 5 serving providers, each with its own pricing and model ID. DeepInfra is currently the cheapest.

What can GLM 5.1 do?

GLM 5.1 supports tool use, reasoning, prompt caching and structured output.

Other Zhipu models

compare pricing across the Zhipu lineup

Popular comparisons

head-to-head pages featuring GLM 5.1
Pricing verified against LiteLLM + OpenRouter + provider pages · refreshed daily

Every weekday

AI moves fast. Here's your debrief.

News, analysis, tools, and more.

For people who build with AI