Z.ai

Z.ai: GLM 4.7 Flash

z-ai/glm-4.7-flash

As a 30B-class SOTA model, GLM-4.7-Flash offers a new option that balances performance and efficiency.

  • Context window: 202,752 tokens
  • Input: text
  • Output: text
  • Pricing: $0.06/M input tokens, $0.4/M output tokens

RuntimeWire coverage

  • Infini-AI-Lab says Vortex hits 3.46x throughput with agent-generated attention

View on OpenRouter. Model data sourced from OpenRouter.

More models from Z.ai

  • Z.ai: GLM 5.3 Flash
  • Z.ai: GLM 5.3 Flash (batch)
  • Z.ai: GLM 5.3
  • Z.ai: GLM 5.2 (free)
  • Z.ai: GLM 5.2
  • Z.ai: GLM 5.1

Browse all AI models