GLM 5.1 vs Grok 4 20 Non Reasoning

API pricing, context window and capabilities for GLM 5.1 and Grok 4 20 Non Reasoning, side by side. Prices are the cheapest serving route per 1M tokens, verified against LiteLLM and OpenRouter.

Verdict

computed from the published numbers

GLM 5.1 wins on price and on intelligence. Grok 4 20 Non Reasoning wins on context window. Benchmark data is unavailable for one or both models.

Price
GLM 5.1
Coding
Not called
Intelligence
GLM 5.1
Context window
Grok 4 20 Non Reasoning

Key takeaways

who wins on what, from the published numbers

GLM 5.1 wins

  • Cheaper input ($1.05 vs $2.00 /1M)
  • Cheaper output ($3.50 vs $6.00 /1M)
  • More serving providers (5 vs 1)
  • Has reasoning mode
  • Supports prompt caching

Grok 4 20 Non Reasoning wins

  • Larger context window (205K vs 2M)
  • Higher max output (131K vs 2M)
  • Supports vision
  • Has web search
Cheaper overall
GLM 5.1
$1.05 + $3.50 /1M
Larger context
Grok 4 20 Non Reasoning
2M
More providers
GLM 5.1
5 providers

Pricing & specs

cheapest serving route · per 1M tokens
GLM 5.1Grok 4 20 Non Reasoning
Input $/1M$1.05 best$2.00
Output $/1M$3.50 best$6.00
Context window205K 2M best
Max output131K 2M best
Serving providers5 best1
Released2026-04-07 2026-03-09

Benchmarks

Artificial Analysis indices, side by side
GLM 5.1Grok 4 20 Non Reasoning
Intelligence41.0 best22.2
Coding55.8 Not available
GPQA Diamond87% best78%
Output speed56 t/s 91 t/s best
Latency (TTFT)1.71s 0.63s best

Cost calculator

choose a serving route for each side
Input tokens / month
Output tokens / month
GLM 5.1
via DeepInfra
$8.75
Grok 4 20 Non Reasoning
via Vertex AI
$16

Capabilities

feature support, side by side
FeatureGLM 5.1Grok 4 20 Non Reasoning
Vision
Tool use
Reasoning
Audio
Prompt caching
Structured output
Web search

Related comparisons

Frequently asked questions

Which is cheaper, GLM 5.1 or Grok 4 20 Non Reasoning?
GLM 5.1 is cheaper for input ($1.05 vs $2.00 per 1M tokens). GLM 5.1 is cheaper for output ($3.50 vs $6.00 per 1M tokens).
What is the context window difference between GLM 5.1 and Grok 4 20 Non Reasoning?
GLM 5.1 has a 205K-token context window and Grok 4 20 Non Reasoning has 2M tokens, so Grok 4 20 Non Reasoning can take a larger prompt.
How do GLM 5.1 and Grok 4 20 Non Reasoning compare on benchmarks?
On the tracked Artificial Analysis indices, GLM 5.1 leads on intelligence.
Pricing verified against LiteLLM + OpenRouter + provider pages · refreshed daily

Every weekday

AI moves fast. Here's your debrief.

News, analysis, tools, and more.

For people who build with AI