z-ai
Z.ai: GLM 5.3 FlashX
z-ai/glm-5.3-flashx
GLM-5.3-FlashX is the high-speed variant of Z.ai's GLM-5.3-Flash, a native multimodal model delivering inference speeds of up to 200 tokens/s.
- Context window: 1,048,576 tokens
- Input: text, image, video
- Output: text
- Pricing: $0.37/M input tokens, $1.25/M output tokens
View on OpenRouter. Model data sourced from OpenRouter.