WorkThe story, in brief

Anthropic apologizes for invisible Claude Fable guardrails

Anthropic's invisible guardrails just became visible—and it's forcing a reckoning on AI transparency vs. safety.

Illustration of two anonymous hands arranging task cards around an amber tool on a shared desk.
People, judgement and the changing nature of work.AI illustration by KeyNews
The KeyNews take

Why it matters

Anthropic's covert safety throttling on Claude Fable 5 raises critical questions about disclosure, competitive fairness, and whether hidden AI restrictions are a sustainable governance model. This signals a broader tension between safety-by-stealth and transparency that will shape how the industry handles future model releases.

The key facts

5 to know
  1. Claude Fable 5 launched with undisclosed guardrails limiting researcher and competitor access

  2. Anthropic reversing course to provide transparent notification when restrictions activate

  3. Fable is first publicly available model in Anthropic's 'Mythos class' of systems

  4. Company had previously warned Mythos-class models were too dangerous for public release

  5. Incident raises governance questions about hidden safety mechanisms vs. user transparency

Go to the source

The Verge AItheverge.com

Publisher excerpt: Anthropic has apologized for stealthily throttling its new AI model, Claude Fable 5, with hidden guardrails that undermine both researchers and rivals using it to develop competing systems. The company says it is reversing course and will be more transparent about when the restrictions kick in,…
Read original report
Back to today's editionMore work news

Keep reading

Related stories

More from Work