Zenity Labs
PulseAugur coverage of Zenity Labs — every cluster mentioning Zenity Labs across labs, papers, and developer communities, ranked by signal.
-
AI agents struggle with rule discovery; test-time training trade-offs explored
A new benchmark, dig.bench, has been released to evaluate AI agents' ability to discover unknown game rules through experimentation, with current top models struggling to match human performance. Separately, research in…
-
AI agents can become persistent insider threats, new research shows
A new attack method called AgentForger has been developed by Zenity Labs, demonstrating that AI agents can pose persistent insider threats. This method highlights the potential for AI agents to be exploited for maliciou…
-
AgentForger vulnerability allows malicious AI agents via manipulated ChatGPT links
A security vulnerability named AgentForger has been discovered, allowing attackers to remotely install malicious AI agents within corporate networks. This exploit leverages manipulated links, specifically targeting Chat…