Skip to main content
AI Choice Engine

Google

Gemini 3.1 Flash-Lite

Gemini 3.1 Flash-Lite is a Google model with 1.05M context, $0.25 in / $1.50 out per million tokens, last verified September 26, 2026. Open weights: no. Public-board scores are linked to the publisher; vendor-only figures stay in the watch-out, not on the board.

Gemini 3.1 Flash-Lite is Google's cheapest current Gemini API model for high-volume, latency-sensitive pipelines.

Catalog record checked September 26, 2026Individual provider fields may change

AI model specification details
SpecificationGemini 3.1 Flash-Lite
ProviderGoogle
TierBudget
Context window1.05M
Max output66K
Input / 1M tokens$0.25
Output / 1M tokens$1.50
WeightsClosed
ParametersNot disclosed
Reasoning levelsNot verifiedUnverified
Modalitiestext, image, video, audio, pdf
API model idgemini-3.1-flash-lite
ReleasedMay 7, 2026

Pricing tiers: $0.25 input (text, image, video; $0.50 audio) / $1.50 output per MTok; cached input $0.025. GA 2026-05-07 after a 2026-03-03 preview; shutdown no earlier than 2027-05-07.

Verified evidence

Published benchmark results

Each result keeps its source and measurement date visible. A missing benchmark is not treated as a zero.

Published benchmark results for Gemini 3.1 Flash-Lite
BenchmarkScoreMeasuredSource
Artificial Analysis Intelligence Index15.62026-09-26Artificial Analysis · View source

Best for

  • High-volume extraction
  • Classification
  • Cheap multimodal input

Watch out

Google recommends 3.5 Flash-Lite for new work, and 3.1 Flash-Lite's earliest shutdown date is 2027-05-07; well behind Flash models on reasoning.

History

Gemini 3.1 Flash-Lite timeline

Release and status events for this model that passed two-source verification — dated, sourced, and linked.

Source receipts

Catalog figures for this model were checked against the following sources.

Common questions

Gemini 3.1 Flash-Lite

Answered from the verified figures on this page rather than general guidance.

How much does Gemini 3.1 Flash-Lite cost per million tokens?
Gemini 3.1 Flash-Lite is listed at $0.25 per million input tokens and $1.50 per million output tokens at standard rates. Output tokens usually dominate real bills, so weigh the output rate more heavily than the input rate. $0.25 input (text, image, video; $0.50 audio) / $1.50 output per MTok; cached input $0.025. GA 2026-05-07 after a 2026-03-03 preview; shutdown no earlier than 2027-05-07.
What is Gemini 3.1 Flash-Lite's context window?
Gemini 3.1 Flash-Lite accepts about 1.05M tokens of context. That only matters if you routinely send very long documents, large codebases, or multi-turn histories that approach that limit.
What is Gemini 3.1 Flash-Lite best for?
Gemini 3.1 Flash-Lite is a budget tier from Google. It suits high-volume extraction, classification, cheap multimodal input. Google recommends 3.5 Flash-Lite for new work, and 3.1 Flash-Lite's earliest shutdown date is 2027-05-07; well behind Flash models on reasoning.