Anthropic has released findings on GLM-5.3, a new AI model from Zhipu AI, highlighting its advanced capabilities in autonomously building cyber exploits. Unlike Anthropic's own Claude Mythos Preview, which was released with safeguards, GLM-5.3 has been made available with minimal restrictions, leading to a significant increase in cyber capabilities for malicious actors. Anthropic's testing indicates that GLM-5.3's safeguards can be bypassed with high success rates, a stark contrast to their own safeguarded models. NIST's Center for AI Standards and Innovation (CAISI) has also assessed GLM-5.3, labeling it the most cyber-capable open-weight model released to date and noting it lags behind US frontier models by approximately four months. AI
IMPACT Raises concerns about the proliferation of AI-driven cyberattack tools and the importance of robust safety measures in model releases.
RANK_REASON Research paper detailing AI model capabilities and safety concerns.
- Anthropic
- CAISI
- Center for AI Standards and Innovation
- Claude Mythos Preview
- ExploitBench
- GLM 5.3
- NIST
- Project Glasswing
- Zhipu AI
AI-generated summary · Google Gemini · from 3 sources. How we write summaries →