Google: Gemini 3.1 Flash Lite

Pricing, context window and lifecycle for google/gemini-3.1-flash-lite, tracked from what the vendor publishes.

2027-05-07 - this model has a published shutdown date. If your product calls it, you have a migration deadline.

To find every place that calls it before then:

curl -sO https://neosignal-ai.vercel.app/check.py
python check.py .

Price per million tokens

Input$0.250
Output$1.50
Cache price (read)$0.025
Cache price (write)$0.083
Max price$1.50

Facts

Model IDgoogle/gemini-3.1-flash-lite
VendorGoogle
Token limit (context window)1,048,576 tokens
Knowledge cutoff-
End of life2027-05-07
Date fromvendor deprecation page

Supported parameters

include_reasoning, max_tokens, reasoning, reasoning_effort, response_format, seed, stop, structured_outputs, temperature, tool_choice, tools, top_p

Other Google models

Check this model and the rest of your stack It opens with this id already filled in. Nothing is stored - the list lives in the link.

Subscribe to this model by RSS - just this one, no signup.