Pricing, context window and lifecycle for z-ai/glm-4.7-flash, tracked from what the vendor publishes.
2026-09-10 - this model has a published shutdown date. If your product calls it, you have a migration deadline.
To find every place that calls it before then:
curl -sO https://neosignal-ai.vercel.app/check.py
python check.py .
| Input | $0.060 |
| Output | $0.400 |
| Cache price (read) | $0.010 |
| Max price | $0.400 |
| Model ID | z-ai/glm-4.7-flash |
| Vendor | Z.ai |
| Token limit (context window) | 202,752 tokens |
| Knowledge cutoff | - |
| End of life | 2026-09-10 |
| Date from | catalogue entry |
frequency_penalty, include_reasoning, logit_bias, logprobs, max_tokens, min_p, presence_penalty, reasoning, repetition_penalty, response_format, seed, stop, structured_outputs, temperature, tool_choice, tools, top_k, top_logprobs, top_p
Every change we have observed since tracking began.
| Detected | Change | Detail |
|---|---|---|
| 2026-09-04 19:14 | Deprecation Set | shutdown date 2026-09-10 announced |
Check this model and the rest of your stack It opens with this id already filled in. Nothing is stored - the list lives in the link.
Subscribe to this model by RSS - just this one, no signup.