SFAI / MODEL PROFILE
Z.ai: GLM 5.3 FlashX (z-ai/glm-5.3-flashx)
GLM-5.3-FlashX is the high-speed variant of Z.ai's GLM-5.3-Flash, a native multimodal model delivering inference speeds of up to 200 tokens/s. Built on the same hybrid sparse and linear attention architecture...
- Developer
- z-ai
- Input modalities
- text, image, video
- Output modalities
- text
- Context tokens
- 1,048,576
- Latest price observation
- 2026-09-28
- Lowest paired standard token quote
- $0.37 input / $1.25 output per 1M tokens
Dated price and provider records
| Provider and channel | Input | Output or native rate | Unit and status | Source and observation |
|---|---|---|---|---|
| Automatic routingopenrouter | $0.37 | $1.25 | 1M tokenspriced · USD | openrouter2026-09-28 |
These are dated, non-executable observations. Verify the exact endpoint, price and terms with the source before use. Unknown values are not free.