Context
1M
Max output
66K
Serving providers
3
Cheapest input
$0.30 /1M
Prices are on-demand list rates. Enterprise commitments (committed-use discounts, provisioned throughput, negotiated rates) can be materially lower.

Benchmarks

independent evaluations · source: Artificial Analysis
Speed & latency
Output speed
300 t/s
tokens / second
Time to first token
18.15s
latency
Headline indices
49.3
of 100
26.8
of 100
Individual benchmarks · click a row for effort variants
BenchmarkBest scoreEffortCost / task
CritPt0%default
GDPval1140default
GPQA Diamond84%default
Humanity's Last Exam18%default
Long-Context Reasoning62%default
MMMU-Pro79%default
Omniscience6.9default
SciCode41%default
TerminalBench v2.154%default
τ-bench Banking16%default

Pricing by serving provider

standard tier · per 1M tokens · click a provider to expand
Serving providerInput /1MOutput /1MEndpoints
Vertex AICheapest$0.30$2.501
EndpointStd inStd outCached inBatch in / outModel ID
StandardCheapest$0.30$2.50$0.03$0.15 / $1.25gemini-3.5-flash-lite
GoogleCheapest$0.30$2.501
OpenRouterCheapest$0.30$2.501

Price history

input + output $/1M since we started tracking
Input Output

Price history is accruing. We record a point each day a price changes; the full backfill lands shortly.

Cost calculator

estimate your monthly spend on this model
$260
estimated / month

Model IDs

copy the exact identifier for your platform
gemini-3.5-flash-litegemini/gemini-3.5-flash-litegoogle/gemini-3.5-flash-lite

Frequently asked questions

Gemini 3.5 Flash-Lite pricing, context and availability

How much does Gemini 3.5 Flash-Lite cost?

Gemini 3.5 Flash-Lite costs $0.30 per 1M input tokens and $2.50 per 1M output tokens at its cheapest provider via Vertex AI.

What is the context window of Gemini 3.5 Flash-Lite?

Gemini 3.5 Flash-Lite accepts up to 1M tokens of context and can return up to 66K output tokens.

Which providers serve Gemini 3.5 Flash-Lite?

Gemini 3.5 Flash-Lite is available from 3 serving providers, each with its own pricing and model ID. Vertex AI is currently the cheapest.

What can Gemini 3.5 Flash-Lite do?

Gemini 3.5 Flash-Lite supports image input (vision), tool use, reasoning, audio input, prompt caching, structured output and web search.

Other Google models

compare pricing across the Google lineup
Pricing verified against LiteLLM + OpenRouter + provider pages · refreshed daily

Every weekday

AI moves fast. Here's your debrief.

News, analysis, tools, and more.

For people who build with AI