AI benchmarks are broken. Here’s what we need instead.
AI benchmarks are broken. Here's what we need instead.

Why it matters
The AI industry's reliance on human-vs-machine benchmarks is fundamentally flawed for enterprise decision-making. This critique matters because leaders are betting billions on model capabilities that may not translate to real-world ROI.
The key facts
4 to knowCurrent benchmark paradigm: AI-vs-human comparison on isolated tasks
Critique: Seductive but insufficient for evaluating practical AI deployment
Source: MIT Technology Review (established academic/professional publication)
Published: March 31, 2026 (recent)
Go to the source
MIT Technology Reviewtechnologyreview.com
Publisher excerpt: For decades, artificial intelligence has been evaluated through the question of whether machines outperform humans. From chess to advanced math, from coding to essay writing, the performance of AI models and applications is tested against that of individual humans completing tasks. This framing is…