Context
1M
Max output
66K
Serving providers
2
Cheapest input
$2.00 /1M
Prices are on-demand list rates. Enterprise commitments (committed-use discounts, provisioned throughput, negotiated rates) can be materially lower.
Benchmarks
independent evaluations · source: Artificial AnalysisHeadline indices
Individual benchmarks · click a row for effort variants
| Benchmark | Best score | Effort | Cost / task | |||||||||||||
|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|
| CritPt | 9% | default | — | |||||||||||||
| ||||||||||||||||
| GPQA Diamond | 91% | default | — | |||||||||||||
| ||||||||||||||||
| Humanity's Last Exam | 40% | default | — | |||||||||||||
| ||||||||||||||||
| IFBench | 70% | default | — | |||||||||||||
| ||||||||||||||||
| LiveCodeBench | 92% | default | — | |||||||||||||
| ||||||||||||||||
| Long-Context Reasoning | 73% | default | — | |||||||||||||
| ||||||||||||||||
| MMLU-Pro | 90% | default | — | |||||||||||||
| ||||||||||||||||
| MMMU-Pro | 80% | default | — | |||||||||||||
| Omniscience | 15.3 | default | — | |||||||||||||
| ||||||||||||||||
| Omniscience Accuracy | 0.6 | default | — | |||||||||||||
| ||||||||||||||||
| Omniscience Non Hallucination | 0.1 | low | — | |||||||||||||
| ||||||||||||||||
| SciCode | 56% | default | — | |||||||||||||
| ||||||||||||||||
| TerminalBench Hard | 42% | default | — | |||||||||||||
| ||||||||||||||||
| τ²-bench | 87% | default | — | |||||||||||||
| ||||||||||||||||
Pricing detail
cost beyond the standard rate · source: models.devLong context (>200K)
$4.00 / $18.00
input / output · per 1M
Pricing by serving provider
standard tier · per 1M tokens · click a provider to expandInput pricing runs from $2.00 to $2.50 per 1M tokens across 2 providers, so the dearest route costs 25% more than the cheapest for the same model.
| Serving provider | Input /1M | Output /1M | Endpoints | ||||||||||||||||
|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|
| ReplicateCheapest | $2.00 | $12.00 | 1 | ||||||||||||||||
| |||||||||||||||||||
| Databricks | $2.50 | $15.00 | 1 | ||||||||||||||||
| |||||||||||||||||||
Price history
input + output $/1M since we started tracking Input Output
Cost calculator
estimate your monthly spend on this model$1,360
estimated / month
Model IDs
copy the exact identifier for your platformreplicate/google/gemini-3-prodatabricks/databricks-gemini-3-pro
Frequently asked questions
Gemini 3 Pro pricing, context and availabilityHow much does Gemini 3 Pro cost?
Gemini 3 Pro costs $2.00 per 1M input tokens and $12.00 per 1M output tokens at its cheapest provider via Replicate. Across 2 serving providers, input prices range from $2.00 to $2.50 per 1M tokens.
What is the context window of Gemini 3 Pro?
Gemini 3 Pro accepts up to 1M tokens of context and can return up to 66K output tokens.
Which providers serve Gemini 3 Pro?
Gemini 3 Pro is available from 2 serving providers, each with its own pricing and model ID. Replicate is currently the cheapest.
What can Gemini 3 Pro do?
Gemini 3 Pro supports image input (vision), tool use, prompt caching and structured output.
Other Google models
compare pricing across the Google lineupPopular comparisons
head-to-head pages featuring Gemini 3 Pro Pricing verified against LiteLLM + OpenRouter + provider pages · refreshed daily
Every weekday
AI moves fast. Here's your debrief.
News, analysis, tools, and more.
For people who build with AI