Context
203K
Max output
131K
Serving providers
5
Cheapest input
$0.00 /1M
Prices are on-demand list rates. Enterprise commitments (committed-use discounts, provisioned throughput, negotiated rates) can be materially lower.

Benchmarks

independent evaluations · source: Artificial Analysis and Hugging Face leaderboards
Speed & latency
Output speed
89 t/s
tokens / second
Time to first token
1.35s
latency
Headline indices
Individual benchmarks · click a row for effort variants
BenchmarkBest scoreEffortCost / task
CritPt0%default
EvalComputeProxy54.2default
GPQA Diamond58%default
Humanity's Last Exam8%default
IFBench61%default
Long-Context Reasoning41%default
Omniscience-62.6default
Omniscience Accuracy0.2default
Omniscience Non Hallucination0.1default
SciCode34%default
SWE-bench Verified59%default
TerminalBench Hard22%default
τ²-bench99%default

Pricing by serving provider

standard tier · per 1M tokens · click a provider to expand
Serving providerInput /1MOutput /1MEndpoints
ZaiCheapest$0.00$0.001
ComponentUnitStandardBatchCached
Text cached in/1M tok$0
Text cache write/1M tok$0
Text input/1M tok$0
Text output/1M tok$0
DeepInfra$0.06$0.401
OpenRouter$0.06$0.401
Cloudflare$0.06$0.401
Novita$0.07$0.401

Price history

input + output $/1M since we started tracking
Input Output
$0.000$0.100$0.200$0.300$0.400Jan 29Jun 24Jul 19Aug 28$0.000$0.000

Cost calculator

estimate your monthly spend on this model
$0
estimated / month

Model IDs

copy the exact identifier for your platform
zai/glm-4.7-flashdeepinfra/zai-org/GLM-4.7-Flashz-ai/glm-4.7-flashcloudflare/@cf/zai-org/glm-4.7-flashnovita/zai-org/glm-4.7-flash

Frequently asked questions

GLM 4.7 Flash pricing, context and availability

How much does GLM 4.7 Flash cost?

GLM 4.7 Flash costs $0.00 per 1M input tokens and $0.00 per 1M output tokens at its cheapest provider via Zai. Across 5 serving providers, input prices range from $0.00 to $0.07 per 1M tokens.

What is the context window of GLM 4.7 Flash?

GLM 4.7 Flash accepts up to 203K tokens of context and can return up to 131K output tokens.

Which providers serve GLM 4.7 Flash?

GLM 4.7 Flash is available from 5 serving providers, each with its own pricing and model ID. Zai is currently the cheapest.

What can GLM 4.7 Flash do?

GLM 4.7 Flash supports image input (vision), tool use, reasoning, prompt caching and structured output.

Other Zhipu models

compare pricing across the Zhipu lineup
Pricing verified against LiteLLM + OpenRouter + provider pages · refreshed daily

Every weekday

AI moves fast. Here's your debrief.

News, analysis, tools, and more.

For people who build with AI