Google Deepmind's "AI co-clinician" beats GPT-5.4 in blind doctor tests but still trails experienced physicians
Google DeepMind's AI co-clinician just beat GPT-5.4 in blind doctor tests. It still can't match experienced physicians—and that gap matters.

Why it matters
Google DeepMind's specialized medical AI outperforms OpenAI's flagship model on clinical tasks in controlled benchmarks, signaling domain-specific model tuning as a competitive differentiator—but the human performance gap reveals real-world deployment risk for AI in high-stakes healthcare decisions.
The key facts
6 to knowGoogle DeepMind AI co-clinician beats GPT-5.4 in blind doctor tests
System still trails experienced physicians in performance
Benchmark: medical simulation/evaluation environment
Comparison model: GPT-5.4 (OpenAI)
Application domain: clinical decision support
Implication: ChatGPT voice mode not ready for medical consultations
Go to the source
The Decoderthe-decoder.com
Publisher excerpt: Google Deepmind is building an "AI co-clinician" to help doctors care for patients. The system shows promising results in simulation studies but still trails experienced physicians. The research also shows why ChatGPT's voice mode isn't ready for serious tasks, let alone medical consultations.