GLM-5.3 是 Z.ai 2026-08-14 对 GLM-5 家族的更新,直指编码 Agent、工具使用、自动化与防御性安全工作。重要变化不是新基座架构——保留 GLM-5.2 基座,扩展后训练(更长更真实的 Agent 环境、更难任务、更完整轨迹、更强验证)。"This is a release about making a large model behave better inside real workflows, not merely making it answer isolated prompts better."
| 项目 | 发布日状态 |
|---|---|
| 发布日期 | 2026-08-14 |
| 开发商 | Z.ai(智谱国际) |
| API 模型 ID | glm-5.3 |
| 主要技术变化 | GLM-5.2 基座上扩展后训练 |
| 推理模式 | 思考必需 |
| 推理档位 | low, high, max(文档默认 max) |
| 主要负载 | 编码、终端 Agent、工具使用、自动化、研究、网络防御 |
| 获取 | GLM Coding Plan、ZCode、文档化的编码 Agent 集成 |
| API 定价 | 发布时官方价格页无 GLM-5.3 行(勿用 5.2 价格估算) |
| 开源权重 | 发布约两周后(待安全评估与加固),发布日不可用 |
| 参数 | 沿用 GLM-5.2 基座描述:约 744B 总参 / 40B 激活(非 5.3 新架构声明) |
| 上下文 | 5.2 文档宣称最高 1M;跑分脚注用 300K / 400K / 1M 不同设置 |
| 基准 | GLM-5.3 | GLM-5.2 | Kimi K3 | DS V4 Pro 0813 | Fable 5 | GPT-5.6 Sol |
|---|---|---|---|---|---|---|
| Terminal-Bench 3.0 | 28.3 | 4.6 | 17.4 | n/a | 33.7 | 34.6 |
| DeepSWE 1.1 | 66.9 | 46.2 | 67.5 | 62.7 | 69.7 | 72.7 |
| NL2Repo | 58.0 | 48.9 | 58.0 | 61.1 | n/a | n/a |
| CyberGym | 84.5 | 77.2 | 80.0 | 83.3 | 83.8 | 83.6 |
| Toolathlon Verified | 73.0 | 59.9 | 76.5 | 74.1 | 74.7 | 74.9 |
| AutomationBench 1.0.6 | 48.2 | 26.2 | 46.7 | 43.2 | 46.2 | 45.8 |
| Agents' Last Exam | 28.5 | 23.8 | 27.6 | 25.7 | 23.8 | 28.6 |
| HLE with Tools | 62.5 | 54.7 | 59.8 | 60.0 | 63.9 | 64.5 |
| GDPval-AA v2 | 1769 | 1508 | 1682 | 1590 | 1743 | 1730 |
5.2→5.3 提升:Terminal-Bench 3.0 +23.7、DeepSWE +20.7、AutomationBench +22.0、Toolathlon +13.1、CyberGym +7.3、GDPval +261。
局限性声明:不同基准的 owner、时间预算、上下文、评分规则与 harness 不同;1 分之差不能作为决定性结论;DeepSWE 66.9 与 Kimi K3 67.5 接近,需 run-level 数据。
API 模型 ID:glm-5.3;思考必需;reasoning_effort 三档。
GLM Coding Plan:所有档位(Lite/Pro/Max)包含 GLM-5.3,月费 $18 起;积分不是独立 API token 价格。
Claude Code / OpenCode 接入路径:官方文档提供(Claude Code 可用 glm-5.3[1m] 后缀 + 1M 压缩窗口)。
权重尚未发布:写"open weights"要用将来时;本地部署暂不可行(无官方 checkpoint 与 serving 配方)。
验证提示:团队应针对自己的仓库、权限、工具 schema 与重试策略测试后训练叙事,而非只信发布表格。
"Thinking is required, with reasoning_effort values of low, high, or max; Z.ai documents max as the default."
"Do not copy GLM-5.2 pricing into a GLM-5.3 cost calculator or assume the two are identical."
"Calling GLM-5.3 'open source' on launch day is premature. … The precise description today is API-available, open-weight… 以上为必要节选,完整内容请查看原文。
GLM-5.3