WorkThe story, in brief

Evaluating chain-of-thought monitorability

OpenAI just proved internal reasoning monitoring beats output checking. Here's why that matters for AI safety at scale.

Illustration of two anonymous hands arranging task cards around an amber tool on a shared desk.
People, judgement and the changing nature of work.AI illustration by KeyNews
The KeyNews take

Why it matters

OpenAI's new framework demonstrates that monitoring a model's reasoning process—not just its outputs—is significantly more effective for controlling advanced AI systems. This addresses a critical governance challenge as models become more capable, offering a concrete technical approach to scalable AI safety and oversight.

The key facts

5 to know
  1. OpenAI introduces chain-of-thought monitorability evaluation framework

  2. Framework covers 13 evaluations across 24 environments

  3. Key finding: Internal reasoning monitoring is more effective than output monitoring alone

  4. Research addresses scalable control for increasingly capable AI systems

  5. Published Dec 18, 2025

Go to the source

OpenAI Blogopenai.com

Publisher excerpt: OpenAI introduces a new framework and evaluation suite for chain-of-thought monitorability, covering 13 evaluations across 24 environments. Our findings show that monitoring a model’s internal reasoning is far more effective than monitoring outputs alone, offering a promising path toward scalable…
Read original report
Back to today's editionMore work news

The wider picture

View all
Illustration of two anonymous hands arranging task cards around an amber tool on a shared desk.
AI illustration by KeyNews
Work01

Trump now says he wants to form an ‘AI Force’

A major political signal on AI governance: the administration is positioning itself to accelerate rather than constrain AI development, with formal institutional backing (czar + task force). Practitioners and policy-watchers need to know the regulatory stance is shifting toward facilitation.

The Verge AI
Illustration of two anonymous hands arranging task cards around an amber tool on a shared desk.
AI illustration by KeyNews
Work02

The 'robot relations' department may become reality in workplace of the future

As corporations deploy autonomous systems across operations, workers face real changes to pay, autonomy, and job structure. The organizational and policy implications of managing human-AI work dynamics are becoming immediate workplace issues.

CNBC Technology
Paper-cut illustration of an amber microchip with circuit paths extending into a row of data-center cabinets.
AI illustration by KeyNews
Work03

AI, data center alarms dominate Congressional Black Caucus week in Washington

AI regulation and data-center expansion are now front-and-center in a major political forum, signaling emerging consensus-building around policy that will affect enterprise AI deployment and the communities hosting compute infrastructure.

CNBC Technology