Context
1M
Max output
66K
Serving providers
5
Cheapest input
$0.38 /1M
Prices are on-demand list rates. Enterprise commitments (committed-use discounts, provisioned throughput, negotiated rates) can be materially lower.

Benchmarks

independent evaluations · source: Artificial Analysis
Speed & latency
Output speed
324 t/s
tokens / second
Time to first token
9.52s
latency
Headline indices
76.1
of 100
45.1
of 100
Individual benchmarks · click a row for effort variants
BenchmarkBest scoreEffortCost / task
Aa Analyst Agent0.6default
Automation Bench0.6default
CritPt14%default
GDPval1531.5default
GPQA Diamond95%default
Harvey Lab0.9default
Humanity's Last Exam48%default
Long-Context Reasoning80%default
Mlcr Overall0.1default
MMMU-Pro85%default
Omniscience26.5default
Omniscience Accuracy0.6default
Omniscience Non Hallucination0.4default
SciCode57%default
TerminalBench v2.186%default
τ-bench Banking33%default

Pricing by serving provider

standard tier · per 1M tokens · click a provider to expand

Input pricing runs from $0.38 to $0.75 per 1M tokens across 5 providers, so the dearest route costs 100% more than the cheapest for the same model.

Serving providerInput /1MOutput /1MEndpoints
OpenRouterCheapest$0.38$1.881
EndpointStd inStd outCached inBatch in / outModel ID
StandardCheapest$0.38$1.88$0.04google/gemini-3.7-flash
Vertex AI$0.75$3.751
Google$0.75$3.751
Vertex AI$0.75$3.751
DeepInfra$0.75$3.751

Price history

input + output $/1M since we started tracking
Input Output
$0.000$0.500$1.00$1.50$2.00Aug 14Aug 28$1.88$0.375

Cost calculator

estimate your monthly spend on this model
$225
estimated / month

Model IDs

copy the exact identifier for your platform
google/gemini-3.7-flashvertex_ai/gemini-3.7-flashgemini/gemini-3.7-flashgemini-3.7-flashdeepinfra/google/gemini-3.7-flash

Frequently asked questions

Gemini 3.7 Flash pricing, context and availability

How much does Gemini 3.7 Flash cost?

Gemini 3.7 Flash costs $0.38 per 1M input tokens and $1.88 per 1M output tokens at its cheapest provider via OpenRouter. Across 5 serving providers, input prices range from $0.38 to $0.75 per 1M tokens.

What is the context window of Gemini 3.7 Flash?

Gemini 3.7 Flash accepts up to 1M tokens of context and can return up to 66K output tokens.

Which providers serve Gemini 3.7 Flash?

Gemini 3.7 Flash is available from 5 serving providers, each with its own pricing and model ID. OpenRouter is currently the cheapest.

What can Gemini 3.7 Flash do?

Gemini 3.7 Flash supports image input (vision), tool use, reasoning, audio input, prompt caching, structured output and web search.

Other Google models

compare pricing across the Google lineup
Pricing verified against LiteLLM + OpenRouter + provider pages · refreshed daily

Every weekday

AI moves fast. Here's your debrief.

News, analysis, tools, and more.

For people who build with AI