Skip to main content
GLM-4.5 Flash is a fast, efficient language model from Zhipu. Call glm-4.5-flash through MixRoute using the endpoint shown below.

Key capabilities

  • OpenAI-compatible - Works with the OpenAI SDK by changing base_url
  • Streaming - Real-time token output through SSE
  • Multi-turn conversations - Uses role-based messages

Quick example

Parameters