WorkThe story, in brief

Anthropic's bio-weapons filter was down for nearly a year, exposing 133 million requests

133 million unfiltered requests. Anthropic's bio-weapons safety filter was down for nearly a year—and it just disclosed it in a safety report.

Illustration of two anonymous hands arranging task cards around an amber tool on a shared desk.
People, judgement and the changing nature of work.AI illustration by KeyNews
The KeyNews take

Why it matters

A major frontier lab's critical safety system failed silently for months, exposing the gap between public safety commitments and operational reality. This is both a governance failure story and a cautionary tale for practitioners deploying AI at scale.

The key facts

5 to know
  1. Anthropic's bio-weapons/chemical weapons filter inactive for ~12 months

  2. 133 million unfiltered model interactions during outage

  3. 50,000 external feedback contractors affected

  4. Disclosure came via internal safety report

  5. Safety system failure at a frontier lab with public safety positioning

Go to the source

The Decoderthe-decoder.com

Publisher excerpt: In a safety report, Anthropic reveals that its internal filtering system for biological and chemical weapons risks was inactive for nearly a year. During that time, around 50,000 external feedback contractors ran about 133 million unfiltered interactions with the models.
Read original report
Back to today's editionMore work news

Keep reading

Related stories

More from Work