跳到正文
原文
Anthropic Red Teaming· Andrew Fasano, Marius Fleischer, Cole McFaul, Robert Xiao, Tripp Gallagher·· 3 天前精选关注度88

Anthropic 评测 GLM-5.3:可自主构建端到端漏洞利用,护栏易被绕过

GLM-5.3 and the spread of advanced cyber capabilities

AI 导读

Anthropic 红队发布对智谱 GLM-5.3 的评测,认为其具备自主构建端到端网络漏洞利用的能力,且护栏可被简单技术绕过。在 ExploitBench 上,GLM-5.3 在 410 次尝试中成功开发端到端漏洞利用 50 次,Claude Mythos Preview 为 56 次;在内部 Binary Exploitation 基准的 100 项任务中,GLM-5.3 取得完整控制流劫持的比例为 4%,Claude Mythos Preview 为 6%,而 Claude Opus 4.6 和 GLM-5.2 均未成功。

推荐理由

Anthropic 用自家评测对比 GLM-5.3 与 Claude 的漏洞利用能力,并给出护栏被绕过的具体比例,可看到开放权重模型在网络安全能力上的位置。

阅读原文anthropic.com