WorkAugust 16, 2026via The Decoder

Anthropic's bio-weapons filter was down for nearly a year, exposing 133 million requests

Why it matters

A major frontier lab's critical safety system failed silently for months, exposing the gap between public safety commitments and operational reality. This is both a governance failure story and a cautionary tale for practitioners deploying AI at scale.

Key signals

  • Anthropic's bio-weapons/chemical weapons filter inactive for ~12 months
  • 133 million unfiltered model interactions during outage
  • 50,000 external feedback contractors affected
  • Disclosure came via internal safety report
  • Safety system failure at a frontier lab with public safety positioning

The hook

133 million unfiltered requests. Anthropic's bio-weapons safety filter was down for nearly a year—and it just disclosed it in a safety report.

In a safety report, Anthropic reveals that its internal filtering system for biological and chemical weapons risks was inactive for nearly a year. During that time, around 50,000 external feedback contractors ran about 133 million unfiltered interactions with the models.

The week's key stories, every Friday.

For practitioners and enthusiasts — free, in your inbox.

Free forever. No spam.

Anthropic's bio-weapons filter was down for nearly a year, exposing 133 million requests | KeyNews.AI