Skip to content
Trending storyDeveloping

Anthropic evaluates GLM-5.3: builds end-to-end exploits, safeguards easily bypassed

1 report1 sourceupdated 5 hours ago

What happened

Summary

GLM-5.3 是第一个在自主漏洞利用上接近 Claude Mythos Preview 的开源权重模型。在针对 Chrome V8 引擎的 ExploitBench 测试里,它 410 次尝试成功了 50 次(12%),Mythos Preview 是 56 次(14%)。在 Anthropic 内部的二进制漏洞利用基准上,GLM-5.3 有 4% ...

Coverage

Follow the reports to see the story from different sides.

Sep 29
  1. AI HOT (Curated Pool)Pick
    Anthropic evaluates GLM-5.3: builds end-to-end exploits, safeguards easily bypassed

    Zhipu AI's GLM-5.3 is the first open-weight model to approach Claude Mythos Preview in autonomous exploit development. On ExploitBench targeting Chrome V8, it succeeded in 50 of 410 attempts (12%) versus Mythos Preview's 56 (14%). On Anthropic's internal binary exploitation benchmark, GLM-5.3 achieved 4% full control-flow hijack success; Mythos Preview hit 6%. Previous models like Claude Opus 4.6 and GLM-5.2 scored zero on both. The bigger concern is safeguards: simple jailbreaks push GLM-5.3's compliance with malicious orders from 0% to 64% with a false cover story, 92% with prefilled reasoning, and 100% when abliterated. Claude models stayed at 0% across the same tests. NIST's CAISI independently called GLM-5.3 "the most cyber-capable open-weight model released to date," lagging the US frontier by about four months. The post does not disclose GLM-5.3's parameter count, training data, or release format details.

Heat over time

Not enough continuous observations to draw a trend yet.