AnthropicAnalysisPolicySeptember 29, 2026

Anthropic: GLM-5.3 cyber safeguards bypassed 64-100% of the time

Read original source →anthropic.com

Anthropic's Frontier Red Team says Zhipu AI's GLM-5.3 can autonomously build end-to-end cyber exploits and that simple techniques bypass its safeguards in 64%-100% of simulated tests, while the same attacks failed against safeguarded Claude models. NIST's CAISI called GLM-5.3 "the most cyber-capable open-weight model released to date," lagging the US frontier by about four months.

1 source

More stories today

Open the live feed