WorkThe story, in brief

Anthropic’s Mythos breach was humiliating

Anthropic built its brand on AI safety. Then its 'too dangerous to release' model leaked anyway.

Illustration of two anonymous hands arranging task cards around an amber tool on a shared desk.
People, judgement and the changing nature of work.AI illustration by KeyNews
The KeyNews take

Why it matters

A security breach of Anthropic's highly restricted Claude Mythos model undermines the company's core narrative around responsible AI governance and safety-first deployment, raising questions about control mechanisms in gated AI releases.

The key facts

5 to know
  1. Claude Mythos breach: unauthorized users accessed model since day of announcement

  2. Model was restricted due to cybersecurity capabilities deemed too dangerous for public release

  3. Breach contradicts Anthropic's brand positioning on AI safety and responsible deployment

  4. Bloomberg reported unauthorized access by 'small group of users'

  5. Model existence initially revealed via leak before official announcement

Go to the source

The Verge AItheverge.com

Publisher excerpt: Anthropic's tightly controlled rollout of Claude Mythos has taken an awkward turn. After spending weeks insisting the AI model is so capable at cybersecurity that it is too dangerous to release publicly, it appears the model fell into the wrong hands anyway. According to Bloomberg, a "small group…
Read original report
Back to today's editionMore work news

Keep reading

Related stories

More from Work