z-ai

Z.ai: GLM 5.3 FlashX

z-ai/glm-5.3-flashx

GLM-5.3-FlashX is the high-speed variant of Z.ai's GLM-5.3-Flash, a native multimodal model delivering inference speeds of up to 200 tokens/s.

  • Context window: 1,048,576 tokens
  • Input: text, image, video
  • Output: text
  • Pricing: $0.37/M input tokens, $1.25/M output tokens

View on OpenRouter. Model data sourced from OpenRouter.