This model was retired on 2026-05-25. The pricing below is the last-known rate, kept for migration reference.
Gemini 3.1 Flash Lite PreviewDeprecated
VisionTool useReasoningAudioPrompt cachingStructured outputWeb search
Context
1M
Max output
66K
Serving providers
3
Cheapest input
$0.25 /1M
Prices are on-demand list rates. Enterprise commitments (committed-use discounts, provisioned throughput, negotiated rates) can be materially lower.
Benchmarks
independent evaluations · source: Artificial AnalysisSpeed & latency
Output speed
299 t/s
tokens / second
Time to first token
5.58s
latency
Headline indices
Intelligence Index
25.6
of 100
Coding Index
34.7
of 100
Agentic Index
6.5
of 100
Individual benchmarks · click a row for effort variants
| Benchmark | Best score | Effort | Cost / task | |
|---|---|---|---|---|
| Aa Analyst Agent | 0.1 | default | — | |
| AA-Briefcase | 0% | — | — | |
| APEX Agents | 12% | default | — | |
| Automation Bench | 0.1 | default | — | |
| CritPt | 1% | default | — | |
| GDPval | 648.3 | default | — | |
| GPQA Diamond | 82% | default | — | |
| Harvey Lab | 0.3 | default | — | |
| Humanity's Last Exam | 17% | default | — | |
| IFBench | 77% | default | — | |
| Long-Context Reasoning | 71% | default | — | |
| Mlcr Overall | 0.1 | default | — | |
| MMMU-Pro | 76% | default | — | |
| Omniscience | -16.4 | default | — | |
| Omniscience Accuracy | 0.4 | default | — | |
| Omniscience Non Hallucination | 0.2 | default | — | |
| SciCode | 42% | default | — | |
| TerminalBench Hard | 24% | default | — | |
| TerminalBench v2.1 | 31% | default | — | |
| τ-bench Banking | 10% | default | — | |
| τ²-bench | 31% | default | — |
Pricing by serving provider
last-known · per 1M tokens · click a provider to expand| Serving provider | Input /1M | Output /1M | Endpoints | |||||||||||||||||||||
|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|
| Vertex AI | $0.25 | $1.50 | 1 | |||||||||||||||||||||
| ||||||||||||||||||||||||
| $0.25 | $1.50 | 1 | ||||||||||||||||||||||
| ||||||||||||||||||||||||
| OpenRouter | $0.25 | $1.50 | 1 | |||||||||||||||||||||
| ||||||||||||||||||||||||
Price history
input + output $/1M since we started tracking Input Output
Cost calculator
estimate your monthly spend on this model$170
estimated / month
Model IDs
copy the exact identifier for your platformgemini-3.1-flash-lite-previewgemini/gemini-3.1-flash-lite-previewgoogle/gemini-3.1-flash-lite-preview
Frequently asked questions
Gemini 3.1 Flash Lite Preview pricing, context and availabilityIs Gemini 3.1 Flash Lite Preview still available?
Gemini 3.1 Flash Lite Preview was retired on 2026-05-25. The pricing on this page is the last-known rate, kept for migration reference.
How much did Gemini 3.1 Flash Lite Preview cost?
Gemini 3.1 Flash Lite Preview's last-known pricing, before it was retired on 2026-05-25, was $0.25 per 1M input tokens and $1.50 per 1M output tokens.
What was the context window of Gemini 3.1 Flash Lite Preview?
Gemini 3.1 Flash Lite Preview had a 1M token context window and could return up to 66K output tokens.
What could Gemini 3.1 Flash Lite Preview do?
Gemini 3.1 Flash Lite Preview supported image input (vision), tool use, reasoning, audio input, prompt caching, structured output and web search.
Other Google models
compare pricing across the Google lineup Pricing verified against LiteLLM + OpenRouter + provider pages · refreshed daily
Every weekday
AI moves fast. Here's your debrief.
News, analysis, tools, and more.
For people who build with AI