This model was retired on 2026-10-20. The pricing below is the last-known rate, kept for migration reference.

Gemini 2.5 Flash LiteDeprecated

GoogleReleasedJun 17, 2025KnowledgeJan 2025
VisionTool useReasoningPrompt cachingStructured outputWeb search
Context
1M
Max output
66K
Serving providers
4
Cheapest input
$0.07 /1M
Prices are on-demand list rates. Enterprise commitments (committed-use discounts, provisioned throughput, negotiated rates) can be materially lower.

Benchmarks

independent evaluations · source: Artificial Analysis
Speed & latency
Output speed
305 t/s
tokens / second
Time to first token
0.28s
latency
Headline indices
Intelligence Index
11.4
of 100
Individual benchmarks · click a row for effort variants
BenchmarkBest scoreEffortCost / task
CritPt0%default
GPQA Diamond63%reasoning: true
Humanity's Last Exam7%reasoning: true
IFBench50%reasoning: true
LiveCodeBench59%reasoning: true
Long-Context Reasoning56%reasoning: true
Mlcr Overall0.0reasoning: true
MMLU-Pro76%reasoning: true
MMMU-Pro58%reasoning: true
Omniscience-45.6reasoning: true
Omniscience Accuracy0.2reasoning: true
Omniscience Non Hallucination0.2reasoning: true
SciCode19%reasoning: true
TerminalBench Hard5%reasoning: true
τ²-bench19%default

Pricing by serving provider

last-known · per 1M tokens · click a provider to expand
Serving providerInput /1MOutput /1MEndpoints
OCI$0.07$0.301
Vertex AI$0.10$0.401
Google$0.10$0.401
OpenRouter$0.10$0.401

Price history

input + output $/1M since we started tracking
Input Output
Input down 25% since first tracked
$0.000$0.100$0.200$0.300$0.400Jul 24Jan 21Apr 4Jul 19$0.300$0.075

Cost calculator

estimate your monthly spend on this model
$39
estimated / month

Model IDs

copy the exact identifier for your platform
oci/google.gemini-2.5-flash-litegemini-2.5-flash-litegemini/gemini-2.5-flash-litegoogle/gemini-2.5-flash-lite

Frequently asked questions

Gemini 2.5 Flash Lite pricing, context and availability

Is Gemini 2.5 Flash Lite still available?

Gemini 2.5 Flash Lite was retired on 2026-10-20. The pricing on this page is the last-known rate, kept for migration reference.

How much did Gemini 2.5 Flash Lite cost?

Gemini 2.5 Flash Lite's last-known pricing, before it was retired on 2026-10-20, was $0.07 per 1M input tokens and $0.30 per 1M output tokens.

What was the context window of Gemini 2.5 Flash Lite?

Gemini 2.5 Flash Lite had a 1M token context window and could return up to 66K output tokens.

What could Gemini 2.5 Flash Lite do?

Gemini 2.5 Flash Lite supported image input (vision), tool use, reasoning, prompt caching, structured output and web search.

Other Google models

compare pricing across the Google lineup
Pricing verified against LiteLLM + OpenRouter + provider pages · refreshed daily

Every weekday

AI moves fast. Here's your debrief.

News, analysis, tools, and more.

For people who build with AI