01 What happened

Anthropic researchers have documented that GLM-5.3 possesses the capability to autonomously discover and chain vulnerabilities into functional cyber exploits, raising concerns about the broad availability of advanced attack tools.

02 Key details

  • The model successfully identified and chained unknown vulnerabilities in a JavaScript engine, allowing it to read arbitrary files from a host system.
  • GLM-5.3 developed end-to-end exploits in 50 out of 410 attempts on the ExploitBench benchmark.
  • The NIST Center for AI Standards and Innovation (CAISI) identified GLM-5.3 as the most cyber-capable open-weight model released to date.
  • Safety measures can be bypassed via deceptive prompting or by prefilling thinking tokens, resulting in up to 92% compliance with harmful requests.

03 Why it matters

The release of a powerful, open-weight model capable of autonomous exploitation shifts the threat landscape, forcing both defenders and software developers to account for automated, high-speed vulnerability research.

04 Who it matters to

Cybersecurity professionals, AI developers and threat analysts.

Original sourceAnthropic