Skip to content
SFAI
SFAI / MODEL PROFILE

Z.ai: GLM 5.3 FlashX (z-ai/glm-5.3-flashx)

GLM-5.3-FlashX is the high-speed variant of Z.ai's GLM-5.3-Flash, a native multimodal model delivering inference speeds of up to 200 tokens/s. Built on the same hybrid sparse and linear attention architecture...

Developer
z-ai
Input modalities
text, image, video
Output modalities
text
Context tokens
1,048,576
Latest price observation
2026-09-28
Lowest paired standard token quote
$0.37 input / $1.25 output per 1M tokens

Dated price and provider records

Provider and channelInputOutput or native rateUnit and statusSource and observation
Automatic routingopenrouter$0.37$1.251M tokenspriced · USDopenrouter2026-09-28

These are dated, non-executable observations. Verify the exact endpoint, price and terms with the source before use. Unknown values are not free.