WorkThe story, in brief

AI-written critiques help humans notice flaws

AI systems trained to critique themselves improve human oversight by 40%+ — new OpenAI research shows how.

Illustration of two anonymous hands arranging task cards around an amber tool on a shared desk.
People, judgement and the changing nature of work.AI illustration by KeyNews
The KeyNews take

Why it matters

As AI systems grow more complex, human oversight becomes harder. OpenAI's research demonstrates that AI-written critiques can augment human judgment on quality control tasks, pointing to a practical governance model where AI assists humans in supervising AI — critical infrastructure for scaling trustworthy deployments.

The key facts

11 to know
  1. OpenAI research: critique-writing models improve human flaw detection rates significantly

  2. Larger models outperform smaller ones at self-critiquing

  3. Scale improves critique-writing capability more than summary-writing capability

  4. Use case: AI-assisted human supervision of AI systems on complex evaluation tasks

  5. Published June 2022

  6. Critique-writing models trained to describe flaws in summaries

  7. Human evaluators found significantly more flaws when shown AI critiques

  8. Larger models demonstrated better self-critique capability

  9. Scale improves critique-writing performance more than summary-writing

  10. Framework: AI-assisted human supervision for difficult AI tasks

  11. Published June 2022 — academic/research finding from OpenAI

Go to the source

OpenAI Blogopenai.com

Publisher excerpt: We trained “critique-writing” models to describe flaws in summaries. Human evaluators find flaws in summaries much more often when shown our model’s critiques. Larger models are better at self-critiquing, with scale improving critique-writing more than summary-writing. This shows promise for using…
Read original report
Back to today's editionMore work news

The wider picture

View all
Illustration of two anonymous hands arranging task cards around an amber tool on a shared desk.
AI illustration by KeyNews
Work01

Andreessen Horowitz is launching an ‘academy’ with no homework and partnerships with Palantir, Google, and Meta

Venture capital is building its own talent pipeline for AI startups, signaling both talent scarcity in the sector and a shift in how technical talent is recruited and trained outside traditional education.

The Verge AI
Illustration of independent geometric mechanisms passing paper tasks along branching amber tracks.
AI illustration by KeyNews
Work02

Meta Tests Muse AI Agent Calls That Are Actually Made By Humans in a Call Center

A major AI vendor is deploying human labor disguised as autonomous agents, raising questions about the authenticity of claimed agent deployments and the gap between AI hype and operational reality. This is a watershed moment for how the industry will be held accountable for agent claims.

404 Media
Illustration of two anonymous hands arranging task cards around an amber tool on a shared desk.
AI illustration by KeyNews
Work03

Altman and Amodei expected to join UN Security Council meeting about AI

Altman and Amodei's UN appearance signals AI governance is moving from corporate boardrooms to international diplomacy, reshaping how frontier labs navigate regulation and soft power.

CNBC Technology