Z.ai: GLM Flash Latest

Pricing, context window and lifecycle for ~z-ai/glm-flash-latest, tracked from what the vendor publishes.

2098-12-31 - this model has a published shutdown date. If your product calls it, you have a migration deadline.

To find every place that calls it before then:

curl -sO https://neosignal-ai.vercel.app/check.py
python check.py .

Price per million tokens

Input$0.075
Output$0.250
Cache price (read)$0.015
Max price$0.250

Facts

Model ID~z-ai/glm-flash-latest
VendorZ.ai latest-alias
Token limit (context window)1,310,720 tokens
Knowledge cutoff-
End of life2098-12-31
Date fromcatalogue entry

Supported parameters

frequency_penalty, include_reasoning, logit_bias, logprobs, max_tokens, min_p, presence_penalty, reasoning, reasoning_effort, repetition_penalty, response_format, seed, stop, structured_outputs, temperature, tool_choice, tools, top_k, top_logprobs, top_p

Other Z.ai latest-alias models

Change history

Every change we have observed since tracking began.

DetectedChangeDetail
2026-09-02 19:14Model AddedZ.ai: GLM Flash Latest

Check this model and the rest of your stack It opens with this id already filled in. Nothing is stored - the list lives in the link.

Subscribe to this model by RSS - just this one, no signup.