GLM 4.5

ZhipuReleasedJul 28, 2025KnowledgeApr 2025
Context
131K
Max output
131K
Serving providers
5
Cheapest input
$0.40 /1M
Prices are on-demand list rates. Enterprise commitments (committed-use discounts, provisioned throughput, negotiated rates) can be materially lower.

Benchmarks

independent evaluations · source: Artificial Analysis
Headline indices
Individual benchmarks · click a row for effort variants
BenchmarkBest scoreEffortCost / task
CritPt0%default
EvalComputeProxy151.2default
GPQA Diamond78%default
Humanity's Last Exam13%default
IFBench44%default
LiveCodeBench74%default
Long-Context Reasoning52%default
MMLU-Pro84%default
Omniscience-27.4default
Omniscience Accuracy0.3default
Omniscience Non Hallucination0.3default
SciCode35%default
TerminalBench Hard22%default
τ²-bench43%default

Pricing by serving provider

standard tier · per 1M tokens · click a provider to expand

Input pricing runs from $0.40 to $0.60 per 1M tokens across 5 providers, so the dearest route costs 50% more than the cheapest for the same model.

Serving providerInput /1MOutput /1MEndpoints
DeepInfraCheapest$0.40$1.601
ComponentUnitStandardBatchCached
Text input/1M tok$0.4000
Text output/1M tok$1.60
Zai$0.60$2.201
Vercel AI Gateway$0.60$2.201
Novita$0.60$2.201
OpenRouter$0.60$2.201

Price history

input + output $/1M since we started tracking
Input Output
$0.000$0.500$1.00$1.50$2.00Jul 30Oct 4Dec 2Jul 19$1.60$0.400

Cost calculator

estimate your monthly spend on this model
$208
estimated / month

Model IDs

copy the exact identifier for your platform
deepinfra/zai-org/GLM-4.5zai/glm-4.5vercel_ai_gateway/zai/glm-4.5novita/zai-org/glm-4.5z-ai/glm-4.5

Frequently asked questions

GLM 4.5 pricing, context and availability

How much does GLM 4.5 cost?

GLM 4.5 costs $0.40 per 1M input tokens and $1.60 per 1M output tokens at its cheapest provider via DeepInfra. Across 5 serving providers, input prices range from $0.40 to $0.60 per 1M tokens.

What is the context window of GLM 4.5?

GLM 4.5 accepts up to 131K tokens of context and can return up to 131K output tokens.

Which providers serve GLM 4.5?

GLM 4.5 is available from 5 serving providers, each with its own pricing and model ID. DeepInfra is currently the cheapest.

What can GLM 4.5 do?

GLM 4.5 supports tool use and reasoning.

Other Zhipu models

compare pricing across the Zhipu lineup
Pricing verified against LiteLLM + OpenRouter + provider pages · refreshed daily

Every weekday

AI moves fast. Here's your debrief.

News, analysis, tools, and more.

For people who build with AI