WorkThe story, in brief

AI benchmarks are broken. Here’s what we need instead.

AI benchmarks are broken. Here's what we need instead.

Illustration of a transparent lens revealing connected networks across layers of paper.
Exploring the next frontier of AI research.AI illustration by KeyNews
The KeyNews take

Why it matters

The AI industry's reliance on human-vs-machine benchmarks is fundamentally flawed for enterprise decision-making. This critique matters because leaders are betting billions on model capabilities that may not translate to real-world ROI.

The key facts

4 to know
  1. Current benchmark paradigm: AI-vs-human comparison on isolated tasks

  2. Critique: Seductive but insufficient for evaluating practical AI deployment

  3. Source: MIT Technology Review (established academic/professional publication)

  4. Published: March 31, 2026 (recent)

Go to the source

MIT Technology Reviewtechnologyreview.com

Publisher excerpt: For decades, artificial intelligence has been evaluated through the question of whether machines outperform humans. From chess to advanced math, from coding to essay writing, the performance of AI models and applications is tested against that of individual humans completing tasks. This framing is…
Read original report
Back to today's editionMore work news

Keep reading

Related stories

More from Work