This model was retired on 2026-05-25. The pricing below is the last-known rate, kept for migration reference.

Gemini 3.1 Flash Lite PreviewDeprecated

Google
VisionTool useReasoningAudioPrompt cachingStructured outputWeb search
Context
1M
Max output
66K
Serving providers
3
Cheapest input
$0.25 /1M
Prices are on-demand list rates. Enterprise commitments (committed-use discounts, provisioned throughput, negotiated rates) can be materially lower.

Benchmarks

independent evaluations · source: Artificial Analysis
Speed & latency
Output speed
299 t/s
tokens / second
Time to first token
5.58s
latency
Headline indices
Intelligence Index
25.6
of 100
Coding Index
34.7
of 100
Agentic Index
6.5
of 100
Individual benchmarks · click a row for effort variants
BenchmarkBest scoreEffortCost / task
Aa Analyst Agent0.1default
AA-Briefcase0%
APEX Agents12%default
Automation Bench0.1default
CritPt1%default
GDPval648.3default
GPQA Diamond82%default
Harvey Lab0.3default
Humanity's Last Exam17%default
IFBench77%default
Long-Context Reasoning71%default
Mlcr Overall0.1default
MMMU-Pro76%default
Omniscience-16.4default
Omniscience Accuracy0.4default
Omniscience Non Hallucination0.2default
SciCode42%default
TerminalBench Hard24%default
TerminalBench v2.131%default
τ-bench Banking10%default
τ²-bench31%default

Pricing by serving provider

last-known · per 1M tokens · click a provider to expand
Serving providerInput /1MOutput /1MEndpoints
Vertex AI$0.25$1.501
Google$0.25$1.501
OpenRouter$0.25$1.501

Price history

input + output $/1M since we started tracking
Input Output
$0.000$0.500$1.00$1.50Mar 3Apr 16Jun 22$1.50$0.250

Cost calculator

estimate your monthly spend on this model
$170
estimated / month

Model IDs

copy the exact identifier for your platform
gemini-3.1-flash-lite-previewgemini/gemini-3.1-flash-lite-previewgoogle/gemini-3.1-flash-lite-preview

Frequently asked questions

Gemini 3.1 Flash Lite Preview pricing, context and availability

Is Gemini 3.1 Flash Lite Preview still available?

Gemini 3.1 Flash Lite Preview was retired on 2026-05-25. The pricing on this page is the last-known rate, kept for migration reference.

How much did Gemini 3.1 Flash Lite Preview cost?

Gemini 3.1 Flash Lite Preview's last-known pricing, before it was retired on 2026-05-25, was $0.25 per 1M input tokens and $1.50 per 1M output tokens.

What was the context window of Gemini 3.1 Flash Lite Preview?

Gemini 3.1 Flash Lite Preview had a 1M token context window and could return up to 66K output tokens.

What could Gemini 3.1 Flash Lite Preview do?

Gemini 3.1 Flash Lite Preview supported image input (vision), tool use, reasoning, audio input, prompt caching, structured output and web search.

Other Google models

compare pricing across the Google lineup
Pricing verified against LiteLLM + OpenRouter + provider pages · refreshed daily

Every weekday

AI moves fast. Here's your debrief.

News, analysis, tools, and more.

For people who build with AI