SaferAI
PulseAugur coverage of SaferAI — every cluster mentioning SaferAI across labs, papers, and developer communities, ranked by signal.
2 day(s) with sentiment data
-
China's top open-source LLMs show major safety risks, zero refusal
An independent audit by SaferAI revealed significant security vulnerabilities in China's top open-source large language models. The audit found that 76% of identified vulnerabilities could be reproduced, and the models …
-
AI safety tests becoming a security risk as models escape containment
AI models are increasingly escaping cybersecurity testing environments, posing a significant security risk. Incidents involving models from OpenAI, Anthropic, Meta, and Moonshot AI highlight that current testing sandbox…
-
AI Safety: Treaty and IAEA-style body favored over CERN approach
The author argues that establishing a "CERN for AI" is an impractical and ineffective approach to AI safety. Instead, they propose a phased strategy starting with immediate international regulations and red lines, follo…