glm-5.3-flash。
核心能力
- 原生多模态:理解文本、图片、视频与文件
- 1M 上下文:支持长上下文专业工作流
- 高效 MoE:总参数 320B、激活参数 18B
- Agent 工作流:支持编程、工具调用与多步执行
快速示例
- cURL
- Python
Documentation Index
Fetch the complete documentation index at: /llms.txt
Use this file to discover all available pages before exploring further.
GLM-5.3 Flash:面向编程、Agent 和视觉任务的高效原生多模态模型。
glm-5.3-flash。
curl "https://api.mixroute.ai/v1/chat/completions" \
-H "Authorization: Bearer YOUR_API_KEY" \
-H "Content-Type: application/json" \
-d '{
"model": "glm-5.3-flash",
"messages": [
{"role": "user", "content": "审查这份部署方案,找出风险最高的回滚缺口。"}
],
"max_tokens": 256,
"reasoning_effort": "high"
}'
from openai import OpenAI
client = OpenAI(api_key="YOUR_API_KEY", base_url="https://api.mixroute.ai/v1")
response = client.chat.completions.create(
model="glm-5.3-flash",
messages=[{"role": "user", "content": "审查这份部署方案,找出风险最高的回滚缺口。"}],
max_tokens=256,
reasoning_effort="high",
)
print(response.choices[0].message.content)
| 参数 | 类型 | 必填 | 说明 |
|---|---|---|---|
model | string | 是 | 固定为 glm-5.3-flash。 |
messages | array | 是 | 包含 role 与 content 的对话消息。 |
max_tokens | integer | 否 | 最大生成 Token 数。 |
stream | boolean | 否 | 启用 SSE 流式输出。 |
reasoning_effort | string | 否 | 推理强度,支持 low、high 和 max。 |
tools | array | 否 | OpenAI 兼容工具定义。 |