Google: Gemini 3.1 Flash Lite Preview

Pricing, context window and lifecycle for google/gemini-3.1-flash-lite-preview, tracked from what the vendor publishes.

2026-05-25 - its vendor retired this model on that date, and the catalogue still lists it. Anything still calling it is running on borrowed time.

To find every place that calls it:

curl -sO https://neosignal-ai.vercel.app/check.py
python check.py .

Price per million tokens

Input$0.250
Output$1.50
Cache price (read)$0.025
Cache price (write)$0.083
Max price$1.50

Facts

Model IDgoogle/gemini-3.1-flash-lite-preview
VendorGoogle
Token limit (context window)1,048,576 tokens
Knowledge cutoff-
Retired on2026-05-25
Date fromvendor deprecation page

Supported parameters

include_reasoning, max_tokens, reasoning, reasoning_effort, response_format, seed, structured_outputs, temperature, tool_choice, tools, top_p

Other Google models

Check this model and the rest of your stack It opens with this id already filled in. Nothing is stored - the list lives in the link.

Subscribe to this model by RSS - just this one, no signup.