This model was retired on 2026-04-02. The pricing below is the last-known rate, kept for migration reference.
Context
205K
Max output
200K
Serving providers
6
Cheapest input
$0.40 /1M
Prices are on-demand list rates. Enterprise commitments (committed-use discounts, provisioned throughput, negotiated rates) can be materially lower.
Benchmarks
independent evaluations · source: Artificial Analysis and Hugging Face leaderboardsSpeed & latency
Output speed
79 t/s
tokens / second
Time to first token
1.19s
latency
Headline indices
Intelligence Index
34.5
of 100
Coding Index
45.3
of 100
Agentic Index
26.2
of 100
Individual benchmarks · click a row for effort variants
| Benchmark | Best score | Effort | Cost / task | |||||||||||||
|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|
| CritPt | 2% | default | — | |||||||||||||
| ||||||||||||||||
| EvalComputeProxy | 1324.5 | default | — | |||||||||||||
| ||||||||||||||||
| GDPval | 1169.1 | default | — | |||||||||||||
| GPQA Diamond | 86% | default | — | |||||||||||||
| ||||||||||||||||
| Humanity's Last Exam | 27% | default | — | |||||||||||||
| ||||||||||||||||
| IFBench | 68% | default | — | |||||||||||||
| ||||||||||||||||
| LiveCodeBench | 89% | default | — | |||||||||||||
| ||||||||||||||||
| Long-Context Reasoning | 68% | default | — | |||||||||||||
| ||||||||||||||||
| MMLU-Pro | 86% | default | — | |||||||||||||
| ||||||||||||||||
| Omniscience | -36.4 | default | — | |||||||||||||
| ||||||||||||||||
| Omniscience Accuracy | 0.3 | default | — | |||||||||||||
| ||||||||||||||||
| Omniscience Non Hallucination | 0.1 | reasoning: false | — | |||||||||||||
| ||||||||||||||||
| SciCode | 45% | default | — | |||||||||||||
| ||||||||||||||||
| SWE-bench Verified | 74% | default | — | |||||||||||||
| TerminalBench Hard | 32% | default | — | |||||||||||||
| ||||||||||||||||
| TerminalBench v2.1 | 45% | default | — | |||||||||||||
| τ-bench Banking | 12% | default | — | |||||||||||||
| τ²-bench | 96% | default | — | |||||||||||||
| ||||||||||||||||
Pricing by serving provider
last-known · per 1M tokens · click a provider to expand| Serving provider | Input /1M | Output /1M | Endpoints | ||||||||||||||||||||||||||
|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|
| DeepInfra | $0.40 | $1.75 | 1 | ||||||||||||||||||||||||||
| |||||||||||||||||||||||||||||
| OpenRouter | $0.40 | $1.75 | 1 | ||||||||||||||||||||||||||
| |||||||||||||||||||||||||||||
| Together AI | $0.45 | $2.00 | 1 | ||||||||||||||||||||||||||
| |||||||||||||||||||||||||||||
| Baseten | $0.60 | $2.20 | 1 | ||||||||||||||||||||||||||
| |||||||||||||||||||||||||||||
| Zai | $0.60 | $2.20 | 1 | ||||||||||||||||||||||||||
| |||||||||||||||||||||||||||||
| Novita | $0.60 | $2.20 | 1 | ||||||||||||||||||||||||||
| |||||||||||||||||||||||||||||
Price history
input + output $/1M since we started tracking Input Output
Input down 33% since first tracked
Cost calculator
estimate your monthly spend on this model$220
estimated / month
Model IDs
copy the exact identifier for your platformdeepinfra/zai-org/GLM-4.7z-ai/glm-4.7together_ai/zai-org/GLM-4.7baseten/zai-org/GLM-4.7zai/glm-4.7novita/zai-org/glm-4.7
Frequently asked questions
GLM 4.7 pricing, context and availabilityIs GLM 4.7 still available?
GLM 4.7 was retired on 2026-04-02. The pricing on this page is the last-known rate, kept for migration reference.
How much did GLM 4.7 cost?
GLM 4.7's last-known pricing, before it was retired on 2026-04-02, was $0.40 per 1M input tokens and $1.75 per 1M output tokens.
What was the context window of GLM 4.7?
GLM 4.7 had a 205K token context window and could return up to 200K output tokens.
What could GLM 4.7 do?
GLM 4.7 supported image input (vision), tool use, reasoning, prompt caching and structured output.
Other Zhipu models
compare pricing across the Zhipu lineup Pricing verified against LiteLLM + OpenRouter + provider pages · refreshed daily
Every weekday
AI moves fast. Here's your debrief.
News, analysis, tools, and more.
For people who build with AI