Context
1M
Max output
66K
Serving providers
4
Cheapest input
$1.50 /1M
Prices are on-demand list rates. Enterprise commitments (committed-use discounts, provisioned throughput, negotiated rates) can be materially lower.

Benchmarks

independent evaluations · source: Artificial Analysis
Speed & latency
Output speed
232 t/s
tokens / second
Time to first token
19.57s
latency
Headline indices
69.2
of 100
38.7
of 100
Individual benchmarks · click a row for effort variants
BenchmarkBest scoreEffortCost / task
CritPt11%default
GDPval1421default
GPQA Diamond93%default
Humanity's Last Exam38%default
Long-Context Reasoning70%default
MMMU-Pro83%default
Omniscience23.5default
SciCode53%default
TerminalBench v2.178%default
τ-bench Banking25%default

Pricing by serving provider

standard tier · per 1M tokens · click a provider to expand
Serving providerInput /1MOutput /1MEndpoints
Vertex AICheapest$1.50$7.501
EndpointStd inStd outCached inBatch in / outModel ID
StandardCheapest$1.50$7.50$0.15$0.75 / $3.75vertex_ai/gemini-3.6-flash
GoogleCheapest$1.50$7.501
Vertex AICheapest$1.50$7.501
OpenRouterCheapest$1.50$7.501

Price history

input + output $/1M since we started tracking
Input Output

Price history is accruing. We record a point each day a price changes; the full backfill lands shortly.

Cost calculator

estimate your monthly spend on this model
$900
estimated / month

Model IDs

copy the exact identifier for your platform
vertex_ai/gemini-3.6-flashgemini/gemini-3.6-flashgemini-3.6-flashgoogle/gemini-3.6-flash

Frequently asked questions

Gemini 3.6 Flash pricing, context and availability

How much does Gemini 3.6 Flash cost?

Gemini 3.6 Flash costs $1.50 per 1M input tokens and $7.50 per 1M output tokens at its cheapest provider via Vertex AI.

What is the context window of Gemini 3.6 Flash?

Gemini 3.6 Flash accepts up to 1M tokens of context and can return up to 66K output tokens.

Which providers serve Gemini 3.6 Flash?

Gemini 3.6 Flash is available from 4 serving providers, each with its own pricing and model ID. Vertex AI is currently the cheapest.

What can Gemini 3.6 Flash do?

Gemini 3.6 Flash supports image input (vision), tool use, reasoning, audio input, prompt caching, structured output and web search.

Other Google models

compare pricing across the Google lineup
Pricing verified against LiteLLM + OpenRouter + provider pages · refreshed daily

Every weekday

AI moves fast. Here's your debrief.

News, analysis, tools, and more.

For people who build with AI