This model was retired on 2027-05-19. The pricing below is the last-known rate, kept for migration reference.

Gemini 3.5 FlashDeprecated

GoogleReleasedMay 22, 2026KnowledgeJan 2025
VisionTool useReasoningAudioPrompt cachingStructured outputWeb search
Context
1M
Max output
66K
Serving providers
5
Cheapest input
$1.50 /1M
Prices are on-demand list rates. Enterprise commitments (committed-use discounts, provisioned throughput, negotiated rates) can be materially lower.

Benchmarks

independent evaluations · source: ARC Prize and Artificial Analysis
Speed & latency
Output speed
204 t/s
tokens / second
Time to first token
19.87s
latency
Headline indices
Intelligence Index
52.0
of 100
Coding Index
70.1
of 100
Agentic Index
39.7
of 100
Individual benchmarks · click a row for effort variants
BenchmarkBest scoreEffortCost / task
Aa Analyst Agent0.5default
AA-Briefcase18%
APEX Agents47%default
ARC-AGI-272%high$0.850
Automation Bench0.4default
CritPt13%default
GDPval1343.4default
GPQA Diamond92%default
Harvey Lab0.8default
Humanity's Last Exam43%default
IFBench76%default
IT-Bench SRE40%default
Long-Context Reasoning81%default
Mlcr Overall0.2default
MMMU-Pro84%default
Omniscience21.2default
Omniscience Accuracy0.5default
Omniscience Non Hallucination0.4medium
SciCode53%default
TerminalBench Hard46%minimal
TerminalBench v2.179%default
τ-bench Banking32%default
τ²-bench96%medium

Pricing detail

cost beyond the standard rate · source: models.dev
Audio
$1.50 / n/a
input / output · per 1M

Pricing by serving provider

last-known · per 1M tokens · click a provider to expand
Serving providerInput /1MOutput /1MEndpoints
Vertex AI$1.50$9.001
Google$1.50$9.001
Vertex AI$1.50$9.001
DeepInfra$1.50$9.001
OpenRouter$1.50$9.001

Price history

input + output $/1M since we started tracking
Input Output
$0.000$5.00$10.00May 19Jun 22Jul 19Aug 21Aug 28$9.00$1.50

Cost calculator

estimate your monthly spend on this model
$1,020
estimated / month

Model IDs

copy the exact identifier for your platform
vertex_ai/gemini-3.5-flashgemini/gemini-3.5-flashgemini-3.5-flashdeepinfra/google/gemini-3.5-flashgoogle/gemini-3.5-flash

Frequently asked questions

Gemini 3.5 Flash pricing, context and availability

Is Gemini 3.5 Flash still available?

Gemini 3.5 Flash was retired on 2027-05-19. The pricing on this page is the last-known rate, kept for migration reference.

How much did Gemini 3.5 Flash cost?

Gemini 3.5 Flash's last-known pricing, before it was retired on 2027-05-19, was $1.50 per 1M input tokens and $9.00 per 1M output tokens.

What was the context window of Gemini 3.5 Flash?

Gemini 3.5 Flash had a 1M token context window and could return up to 66K output tokens.

What could Gemini 3.5 Flash do?

Gemini 3.5 Flash supported image input (vision), tool use, reasoning, audio input, prompt caching, structured output and web search.

Other Google models

compare pricing across the Google lineup

Popular comparisons

head-to-head pages featuring Gemini 3.5 Flash
Pricing verified against LiteLLM + OpenRouter + provider pages · refreshed daily

Every weekday

AI moves fast. Here's your debrief.

News, analysis, tools, and more.

For people who build with AI