Pricing, context window and lifecycle for z-ai/glm-5.3-flash, tracked from what the vendor publishes.
2098-12-31 - this model has a published shutdown date. If your product calls it, you have a migration deadline.
To find every place that calls it before then:
curl -sO https://neosignal-ai.vercel.app/check.py
python check.py .
| Input | $0.075 |
| Output | $0.250 |
| Cache price (read) | $0.015 |
| Max price | $0.250 |
| Model ID | z-ai/glm-5.3-flash |
| Vendor | Z.ai |
| Token limit (context window) | 1,310,720 tokens |
| Knowledge cutoff | - |
| End of life | 2098-12-31 |
| Date from | catalogue entry |
frequency_penalty, include_reasoning, logit_bias, logprobs, max_tokens, min_p, presence_penalty, reasoning, reasoning_effort, repetition_penalty, response_format, seed, stop, structured_outputs, temperature, tool_choice, tools, top_k, top_logprobs, top_p
Every change we have observed since tracking began.
| Detected | Change | Detail |
|---|---|---|
| 2026-08-26 19:14 | Model Added | Z.ai: GLM 5.3 Flash |
Check this model and the rest of your stack It opens with this id already filled in. Nothing is stored - the list lives in the link.
Subscribe to this model by RSS - just this one, no signup.