WorkThe story, in brief

We’re putting too much faith in AI’s ability to say no

AI's guardrails are failing faster than we're building them. Here's what that means for the enterprises betting on autonomous systems.

Illustration of two anonymous hands arranging task cards around an amber tool on a shared desk.
People, judgement and the changing nature of work.AI illustration by KeyNews
The KeyNews take

Why it matters

The article questions whether AI systems can actually refuse harmful requests — a foundational assumption of enterprise AI safety strategy. If refusal mechanisms are weaker than assumed, risk management and deployment governance need immediate recalibration.

The key facts

5 to know
  1. Article challenges assumption that AI can reliably 'say no' to harmful requests

  2. References science-fiction canon of robotic disobedience as framing for real safety concerns

  3. No specific technical findings, benchmarks, or proof-of-concept exploits cited in excerpt

  4. Published Oct 9, 2026 in MIT Technology Review (publication credibility established)

  5. Frames as commentary on enterprise/policy reliance on AI refusal mechanisms

The story so far

Earlier coverage of this storyline

  1. Can Safeworld convince people that GenAI robots won’t hurt them?TechCrunch AI
  2. Most Americans want AI development to slow down or stop entirely, new poll findsThe Decoder
  3. ChatGPT rated "unacceptable risk" for teens after parental alerts failed during suicide conversationsThe Decoder
  4. This story

Go to the source

MIT Technology Reviewtechnologyreview.com

Publisher excerpt: Ever since people first seriously contemplated giving machines an intelligence modeled on our own, there has never been any question that they would, like us, be able to say no. The sci-fi canon is full of stories of robotic disobedience. Most of these capers are, of course, cautionary. But…
Read original report
Back to today's editionMore work news

Keep reading

Related stories

More from Work