Z.ai: GLM 5.3 Flash

Pricing, context window and lifecycle for z-ai/glm-5.3-flash, tracked from what the vendor publishes.

2098-12-31 - this model has a published shutdown date. If your product calls it, you have a migration deadline.

To find every place that calls it before then:

curl -sO https://neosignal-ai.vercel.app/check.py
python check.py .

Price per million tokens

Input$0.075
Output$0.250
Cache price (read)$0.015
Max price$0.250

Facts

Model IDz-ai/glm-5.3-flash
VendorZ.ai
Token limit (context window)1,310,720 tokens
Knowledge cutoff-
End of life2098-12-31
Date fromcatalogue entry

Supported parameters

frequency_penalty, include_reasoning, logit_bias, logprobs, max_tokens, min_p, presence_penalty, reasoning, reasoning_effort, repetition_penalty, response_format, seed, stop, structured_outputs, temperature, tool_choice, tools, top_k, top_logprobs, top_p

Other Z.ai models

Change history

Every change we have observed since tracking began.

DetectedChangeDetail
2026-08-26 19:14Model AddedZ.ai: GLM 5.3 Flash

Check this model and the rest of your stack It opens with this id already filled in. Nothing is stored - the list lives in the link.

Subscribe to this model by RSS - just this one, no signup.