Z.ai: GLM 4.7 Flash

Pricing, context window and lifecycle for z-ai/glm-4.7-flash, tracked from what the vendor publishes.

2026-09-10 - this model has a published shutdown date. If your product calls it, you have a migration deadline.

To find every place that calls it before then:

curl -sO https://neosignal-ai.vercel.app/check.py
python check.py .

Price per million tokens

Input$0.060
Output$0.400
Cache price (read)$0.010
Max price$0.400

Facts

Model IDz-ai/glm-4.7-flash
VendorZ.ai
Token limit (context window)202,752 tokens
Knowledge cutoff-
End of life2026-09-10
Date fromcatalogue entry

Supported parameters

frequency_penalty, include_reasoning, logit_bias, logprobs, max_tokens, min_p, presence_penalty, reasoning, repetition_penalty, response_format, seed, stop, structured_outputs, temperature, tool_choice, tools, top_k, top_logprobs, top_p

Other Z.ai models

Change history

Every change we have observed since tracking began.

DetectedChangeDetail
2026-09-04 19:14Deprecation Setshutdown date 2026-09-10 announced

Check this model and the rest of your stack It opens with this id already filled in. Nothing is stored - the list lives in the link.

Subscribe to this model by RSS - just this one, no signup.