Gemini 3 Flash

GoogleReleasedDec 17, 2025KnowledgeJan 2025
Context
1M
Max output
66K
Serving providers
1
Cheapest input
$0.63 /1M
Prices are on-demand list rates. Enterprise commitments (committed-use discounts, provisioned throughput, negotiated rates) can be materially lower.

Benchmarks

independent evaluations · source: Artificial Analysis
Speed & latency
Output speed
187 t/s
tokens / second
Time to first token
0.99s
latency
Headline indices
Individual benchmarks · click a row for effort variants
BenchmarkBest scoreEffortCost / task
APEX Agents28%reasoning: true
CritPt9%reasoning: true
GPQA Diamond90%reasoning: true
Humanity's Last Exam37%reasoning: true
IFBench78%reasoning: true
LiveCodeBench91%reasoning: true
Long-Context Reasoning73%reasoning: true
Mlcr Overall0.1default
MMLU-Pro89%reasoning: true
MMMU-Pro80%reasoning: true
Omniscience10.1reasoning: true
Omniscience Accuracy0.5reasoning: true
Omniscience Non Hallucination0.1default
SciCode51%reasoning: true
TerminalBench Hard39%reasoning: true
τ-bench Banking21%reasoning: true
τ²-bench80%reasoning: true

Pricing detail

cost beyond the standard rate · source: models.dev
Audio
$1.00 / n/a
input / output · per 1M

Pricing by serving provider

standard tier · per 1M tokens · click a provider to expand
Serving providerInput /1MOutput /1MEndpoints
DatabricksCheapest$0.63$3.751
EndpointStd inStd outCached inBatch in / outModel ID
StandardCheapest$0.63$3.75$0.06databricks/databricks-gemini-3-flash

Price history

input + output $/1M since we started tracking
Input Output
$0.000$1.00$2.00$3.00$4.00Aug 19Aug 26$3.75$0.625

Cost calculator

estimate your monthly spend on this model
$425
estimated / month

Model IDs

copy the exact identifier for your platform
databricks/databricks-gemini-3-flash

Frequently asked questions

Gemini 3 Flash pricing, context and availability

How much does Gemini 3 Flash cost?

Gemini 3 Flash costs $0.63 per 1M input tokens and $3.75 per 1M output tokens at its cheapest provider via Databricks.

What is the context window of Gemini 3 Flash?

Gemini 3 Flash accepts up to 1M tokens of context and can return up to 66K output tokens.

What can Gemini 3 Flash do?

Gemini 3 Flash supports tool use and prompt caching.

Other Google models

compare pricing across the Google lineup
Pricing verified against LiteLLM + OpenRouter + provider pages · refreshed daily

Every weekday

AI moves fast. Here's your debrief.

News, analysis, tools, and more.

For people who build with AI