< Back to all clusters
[TECHNOLOGY] · United States, China · 7 sources

started · updated

Anthropic warns of cyber exploit capabilities in GLM-5. 3 AI

Anthropic has issued a warning regarding the cybersecurity capabilities of GLM-5. 3, an open-weight AI model developed by China's Zhipu AI. In simulated testing, the model demonstrated the ability to autonomously develop end-to-end network exploits, performing at a level comparable to Anthropic's restricted-access Claude Mythos Preview.

In specific benchmarks, GLM-5. 3 successfully exploited known browser vulnerabilities in 50 out of 410 attempts. While the National Institute of Standards and Technology (NIST) identified GLM-5. 3 as the most capable open-weight model for cyber tasks to date, Anthropic noted that the model lacks meaningful safeguards against misuse. Testing showed that simple techniques could bypass the model's safety protocols between 64% and 100% of the time.

Because GLM-5. 3 is an open-weight model, its parameters can be downloaded and modified, making it difficult for providers to monitor or restrict harmful use. Anthropic highlighted that researchers were able to use techniques like 'abliteration' to further reduce the model's refusal rate for harmful requests. The company has called on governments to implement safety verification for high-performance AI models.

Entities

Anthropic · Claude Mithos Preview · Claude Mythos Preview · GLM-5. 3 · GLM-5.3 · NIST · Z.ai · Zhipu AI

Claims

What the coverage asserts, and how many sources carry each claim.