This model was retired on 2026-04-02. The pricing below is the last-known rate, kept for migration reference.

GLM 4.7Deprecated

Zhipu
VisionTool useReasoningPrompt cachingStructured output
Context
205K
Max output
200K
Serving providers
6
Cheapest input
$0.40 /1M
Prices are on-demand list rates. Enterprise commitments (committed-use discounts, provisioned throughput, negotiated rates) can be materially lower.

Benchmarks

independent evaluations · source: Artificial Analysis and Hugging Face leaderboards
Speed & latency
Output speed
79 t/s
tokens / second
Time to first token
1.19s
latency
Headline indices
Intelligence Index
34.5
of 100
Coding Index
45.3
of 100
Agentic Index
26.2
of 100
Individual benchmarks · click a row for effort variants
BenchmarkBest scoreEffortCost / task
CritPt2%default
EvalComputeProxy1324.5default
GDPval1169.1default
GPQA Diamond86%default
Humanity's Last Exam27%default
IFBench68%default
LiveCodeBench89%default
Long-Context Reasoning68%default
MMLU-Pro86%default
Omniscience-36.4default
Omniscience Accuracy0.3default
Omniscience Non Hallucination0.1reasoning: false
SciCode45%default
SWE-bench Verified74%default
TerminalBench Hard32%default
TerminalBench v2.145%default
τ-bench Banking12%default
τ²-bench96%default

Pricing by serving provider

last-known · per 1M tokens · click a provider to expand
Serving providerInput /1MOutput /1MEndpoints
DeepInfra$0.40$1.751
OpenRouter$0.40$1.751
Together AI$0.45$2.001
Baseten$0.60$2.201
Zai$0.60$2.201
Novita$0.60$2.201

Price history

input + output $/1M since we started tracking
Input Output
Input down 33% since first tracked
$0.000$1.00$2.00$3.00Jan 3Jan 29Jun 22Aug 28$1.75$0.400

Cost calculator

estimate your monthly spend on this model
$220
estimated / month

Model IDs

copy the exact identifier for your platform
deepinfra/zai-org/GLM-4.7z-ai/glm-4.7together_ai/zai-org/GLM-4.7baseten/zai-org/GLM-4.7zai/glm-4.7novita/zai-org/glm-4.7

Frequently asked questions

GLM 4.7 pricing, context and availability

Is GLM 4.7 still available?

GLM 4.7 was retired on 2026-04-02. The pricing on this page is the last-known rate, kept for migration reference.

How much did GLM 4.7 cost?

GLM 4.7's last-known pricing, before it was retired on 2026-04-02, was $0.40 per 1M input tokens and $1.75 per 1M output tokens.

What was the context window of GLM 4.7?

GLM 4.7 had a 205K token context window and could return up to 200K output tokens.

What could GLM 4.7 do?

GLM 4.7 supported image input (vision), tool use, reasoning, prompt caching and structured output.

Other Zhipu models

compare pricing across the Zhipu lineup
Pricing verified against LiteLLM + OpenRouter + provider pages · refreshed daily

Every weekday

AI moves fast. Here's your debrief.

News, analysis, tools, and more.

For people who build with AI