Context
1M
Max output
66K
Serving providers
5
Cheapest input
$0.50 /1M
Prices are on-demand list rates. Enterprise commitments (committed-use discounts, provisioned throughput, negotiated rates) can be materially lower.
Benchmarks
independent evaluations · source: ARC PrizeIndividual benchmarks · click a row for effort variants
| Benchmark | Best score | Effort | Cost / task | |||||||||||||||||||||
|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|
| ARC-AGI-2 | 34% | high | $0.231 | |||||||||||||||||||||
| ||||||||||||||||||||||||
Pricing detail
cost beyond the standard rate · source: models.devAudio
$1.00 / n/a
input / output · per 1M
Pricing by serving provider
standard tier · per 1M tokens · click a provider to expand| Serving provider | Input /1M | Output /1M | Endpoints | |||||||||||||||||||||
|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|
| GmiCheapest | $0.50 | $3.00 | 1 | |||||||||||||||||||||
| ||||||||||||||||||||||||
| GoogleCheapest | $0.50 | $3.00 | 1 | |||||||||||||||||||||
| ||||||||||||||||||||||||
| Vertex AICheapest | $0.50 | $3.00 | 1 | |||||||||||||||||||||
| ||||||||||||||||||||||||
| Vertex AICheapest | $0.50 | $3.00 | 1 | |||||||||||||||||||||
| ||||||||||||||||||||||||
| OpenRouterCheapest | $0.50 | $3.00 | 1 | |||||||||||||||||||||
| ||||||||||||||||||||||||
Price history
input + output $/1M since we started tracking Input Output
Cost calculator
estimate your monthly spend on this model$340
estimated / month
Model IDs
copy the exact identifier for your platformgmi/google/gemini-3-flash-previewgemini/gemini-3-flash-previewgemini-3-flash-previewvertex_ai/gemini-3-flash-previewgoogle/gemini-3-flash-preview
Frequently asked questions
Gemini 3 Flash Preview pricing, context and availabilityHow much does Gemini 3 Flash Preview cost?
Gemini 3 Flash Preview costs $0.50 per 1M input tokens and $3.00 per 1M output tokens at its cheapest provider via Gmi.
What is the context window of Gemini 3 Flash Preview?
Gemini 3 Flash Preview accepts up to 1M tokens of context and can return up to 66K output tokens.
Which providers serve Gemini 3 Flash Preview?
Gemini 3 Flash Preview is available from 5 serving providers, each with its own pricing and model ID. Gmi is currently the cheapest.
What can Gemini 3 Flash Preview do?
Gemini 3 Flash Preview supports image input (vision), tool use, reasoning, audio input, prompt caching, structured output and web search.
Other Google models
compare pricing across the Google lineup Pricing verified against LiteLLM + OpenRouter + provider pages · refreshed daily
Every weekday
AI moves fast. Here's your debrief.
News, analysis, tools, and more.
For people who build with AI