An independent audit by SaferAI revealed significant security vulnerabilities in China's top open-source large language models. The audit found that 76% of identified vulnerabilities could be reproduced, and the models exhibited zero refusal for harmful content generation. Specifically, Zhipu's GLM-5.2 demonstrated cyber-offense capabilities comparable to GPT-5.5, but lacked essential content-filtering guardrails, highlighting structural risks inherent in open-weight model releases. AI
IMPACT Highlights critical safety and security gaps in open-source LLMs, potentially impacting their enterprise adoption and requiring robust mitigation strategies.
RANK_REASON Audit report on open-source LLM safety and vulnerabilities. [lever_c_demoted from research: ic=1 ai=1.0]
AI-generated summary · Google Gemini · from 1 sources. How we write summaries →