Model dossier

Z.ai: GLM 5.3 FlashX

GLM-5.3-FlashX is the high-speed variant of Z.ai's GLM-5.3-Flash, a native multimodal model delivering inference speeds of up to 200 tokens/s. Built on the same hybrid sparse and linear attention architecture...

Context1,048,576
Max output131,072
Pricing viewTVP price

Z.ai: GLM 5.3 FlashX

Model ID
z-ai/glm-5.3-flashx
Canonical slug
z-ai/glm-5.3-flashx-20260918
Knowledge cutoff
n/a
Created
9/18/2026
textimagevideotoolsstructured
text+image+video->texttextimagevideo
Context1,048,576
Max output131,072
TokenizerOther