FrontierThe story, in brief

Google Deepmind's "AI co-clinician" beats GPT-5.4 in blind doctor tests but still trails experienced physicians

Google DeepMind's AI co-clinician just beat GPT-5.4 in blind doctor tests. It still can't match experienced physicians—and that gap matters.

Illustration of a transparent lens revealing connected networks across layers of paper.
Exploring the next frontier of AI research.AI illustration by KeyNews
The KeyNews take

Why it matters

Google DeepMind's specialized medical AI outperforms OpenAI's flagship model on clinical tasks in controlled benchmarks, signaling domain-specific model tuning as a competitive differentiator—but the human performance gap reveals real-world deployment risk for AI in high-stakes healthcare decisions.

The key facts

6 to know
  1. Google DeepMind AI co-clinician beats GPT-5.4 in blind doctor tests

  2. System still trails experienced physicians in performance

  3. Benchmark: medical simulation/evaluation environment

  4. Comparison model: GPT-5.4 (OpenAI)

  5. Application domain: clinical decision support

  6. Implication: ChatGPT voice mode not ready for medical consultations

Go to the source

The Decoderthe-decoder.com

Publisher excerpt: Google Deepmind is building an "AI co-clinician" to help doctors care for patients. The system shows promising results in simulation studies but still trails experienced physicians. The research also shows why ChatGPT's voice mode isn't ready for serious tasks, let alone medical consultations.
Read original report
Back to today's editionMore frontier news

Keep reading

Related stories

More from Frontier