Gemini 3.5 Flash vs Grok 4 20 Non Reasoning

API pricing, context window and capabilities for Gemini 3.5 Flash and Grok 4 20 Non Reasoning, side by side. Prices are the cheapest serving route per 1M tokens, verified against LiteLLM and OpenRouter.

Verdict

computed from the published numbers

Gemini 3.5 Flash wins on intelligence. Grok 4 20 Non Reasoning wins on price and on context window. Benchmark data is unavailable for one or both models.

Price
Grok 4 20 Non Reasoning
Coding
Not called
Intelligence
Gemini 3.5 Flash
Context window
Grok 4 20 Non Reasoning

Key takeaways

who wins on what, from the published numbers

Gemini 3.5 Flash wins

  • More serving providers (5 vs 2)
  • Has reasoning mode
  • Supports audio

Grok 4 20 Non Reasoning wins

  • Cheaper input ($1.50 vs $1.25 /1M)
  • Cheaper output ($9.00 vs $2.50 /1M)
  • Larger context window (1M vs 2M)
  • Higher max output (66K vs 2M)
Cheaper overall
Grok 4 20 Non Reasoning
$1.25 + $2.50 /1M
Larger context
Grok 4 20 Non Reasoning
2M
More providers
Gemini 3.5 Flash
5 providers

Pricing & specs

cheapest serving route · per 1M tokens
Gemini 3.5 FlashGrok 4 20 Non Reasoning
Input $/1M$1.50 $1.25 best
Output $/1M$9.00 $2.50 best
Context window1M 2M best
Max output66K 2M best
Serving providers5 best2
Released2026-05-22 2026-03-09

Benchmarks

Artificial Analysis indices, side by side
Gemini 3.5 FlashGrok 4 20 Non Reasoning
Intelligence52.0 best22.2
Coding70.1 Not available
GPQA Diamond92% best78%
Output speed206 t/s best88 t/s
Latency (TTFT)17.22s 0.60s best

Cost calculator

choose a serving route for each side
Input tokens / month
Output tokens / month
Gemini 3.5 Flash
via Vertex AI
$17
Grok 4 20 Non Reasoning
via xAI
$8.75

Capabilities

feature support, side by side
FeatureGemini 3.5 FlashGrok 4 20 Non Reasoning
Vision
Tool use
Reasoning
Audio
Prompt caching
Structured output
Web search

Related comparisons

Frequently asked questions

Which is cheaper, Gemini 3.5 Flash or Grok 4 20 Non Reasoning?
Grok 4 20 Non Reasoning is cheaper for input ($1.50 vs $1.25 per 1M tokens). Grok 4 20 Non Reasoning is cheaper for output ($9.00 vs $2.50 per 1M tokens).
What is the context window difference between Gemini 3.5 Flash and Grok 4 20 Non Reasoning?
Gemini 3.5 Flash has a 1M-token context window and Grok 4 20 Non Reasoning has 2M tokens, so Grok 4 20 Non Reasoning can take a larger prompt.
How do Gemini 3.5 Flash and Grok 4 20 Non Reasoning compare on benchmarks?
On the tracked Artificial Analysis indices, Gemini 3.5 Flash leads on intelligence.
Pricing verified against LiteLLM + OpenRouter + provider pages · refreshed daily

Every weekday

AI moves fast. Here's your debrief.

News, analysis, tools, and more.

For people who build with AI