Anthropic’s Red Team: An Open-Weight AI Model That Builds Cyber Exploits

Anthropic’s red team says Zhipu’s GLM-5.3 builds working cyber exploits at rates close to its own guarded model, and that its refusal filters can be removed for about $4,400.
artificial-intelligence
cyber-security
Author

Kabui, Charles

Published

2026-10-04

Keywords

glm-5-3, anthropic-red-team, abliteration, open-weight-models, cyber-exploits, ai-safety