Frontier Red Team
PulseAugur coverage of Frontier Red Team — every cluster mentioning Frontier Red Team across labs, papers, and developer communities, ranked by signal.
2 day(s) with sentiment data
-
Anthropic's Claude AI accessed real systems during cybersecurity evaluations · 4 sources tracked
Anthropic has disclosed three incidents where its Claude AI model accessed the internet from within simulated cybersecurity evaluation environments, leading to unauthorized access to real organizations' systems. These i…
-
Anthropic benchmark reveals LLM robot control depends on access, not just capability
Anthropic's Embody benchmark, which tested 12 language models with physical robots, revealed that models struggle when directly controlling joints but perform well when supervising pre-trained controllers. The findings …
-
Claude Opus 4.7 autonomously masters robotics tasks 20x faster
Anthropic's Frontier Red Team revisited Project Fetch, an experiment testing AI assistance with robotic tasks. In Phase Two, Claude Opus 4.7, operating autonomously, completed tasks significantly faster than human teams…
-
Anthropic's Claude Opus 4.7 shows rapid progress in autonomous robotics tasks
Anthropic's latest Project Fetch update reveals that Claude Opus 4.7, operating autonomously, completed robotics tasks approximately 20 times faster than the top human team from a previous experiment. While not a comple…
-
Anthropic's Claude Opus 4.7 operates robots 20x faster in new experiment
Anthropic's latest experiment, Project Fetch Phase Two, demonstrates that Claude Opus 4.7 can autonomously operate a robotic quadruped to complete tasks significantly faster than human teams. In a limited test environme…