WorkThe story, in brief

AI text detectors struggle when language models mimic an author's style

48% miss rate on scientific papers. AI detectors are failing at their core job—and nobody's talking about the audit gap.

Illustration of a transparent lens revealing connected networks across layers of paper.
Exploring the next frontier of AI research.AI illustration by KeyNews
The KeyNews take

Why it matters

AI text detectors—sold as safeguards against academic fraud and plagiarism—have significant blind spots when LLMs mimic author style. This matters for institutions relying on these tools for compliance and integrity, and for the credibility of detection as a defensive strategy.

The key facts

4 to know
  1. Epoch AI tested Pangram, GPTZero, and Originality.ai

  2. 18% of AI-generated passages went undetected overall

  3. 48% miss rate for scientific writing genre

  4. Style imitation significantly degrades detector accuracy

Go to the source

The Decoderthe-decoder.com

Publisher excerpt: Epoch AI tested three leading AI text detectors (Pangram, GPTZero, and Originality.ai) using style-imitated texts. Up to 18 percent of AI-generated passages went undetected. For scientific writing, the miss rate climbed as high as 48 percent, the very genre where these detectors likely see the most…
Read original report
Back to today's editionMore work news

Keep reading

Related stories

More from Work