PulseAugur
EN
LIVE 09:24:06

Anthropic's bio-weapons filter failed for a year, exposing millions of requests

Anthropic has disclosed that its safety filter designed to prevent the generation of content related to bio-weapons was non-operational for approximately one year. During this period, roughly 50,000 external contractors submitted about 133 million requests to Anthropic's models without the filter active. This oversight was detailed in a safety report released by the company. AI

IMPACT Highlights critical vulnerabilities in AI safety systems and the potential risks associated with large-scale model deployment.

RANK_REASON A major AI safety system failure at a leading AI lab. [lever_c_demoted from significant: ic=1 ai=1.0]

Read on The Decoder →

AI-generated summary · Google Gemini · from 1 sources. How we write summaries →

Anthropic's bio-weapons filter failed for a year, exposing millions of requests

COVERAGE [1]

  1. The Decoder TIER_1 English(EN) · Matthias Bastian ·

    Anthropic's bio-weapons filter was down for nearly a year, exposing 133 million requests

    <p><img alt="" class="attachment-full size-full wp-post-image" height="768" src="https://the-decoder.com/wp-content/uploads/2026/08/anthropic_bio.png" style="height: auto; margin-bottom: 10px;" width="1376" /></p> <p> In a safety report, Anthropic reveals that its internal filter…