AIToday
Fortune AIPublished: May 5, 2026, 04:00 JST1 min read

Harvard study finds OpenAI's o1-preview outperforms emergency room physicians at diagnosis

Harvard study finds OpenAI's o1-preview outperforms emergency room physicians at diagnosis

3 Key Points

  1. Researchers at Harvard Medical School and Beth Israel Deaconess Medical Center compared emergency room diagnoses from OpenAI's o1-preview against those from two internal medicine attending physicians. Two other attending physicians, unaware of the source, assessed both sets of diagnoses and favored the AI model.

  2. Unlike prior AI medical studies, researchers presented each case exactly as it appeared in an electronic health record without cleaning up the data. Peter Brodeur, a study co-author, noted that AI models now score consistently close to 100% on multiple-choice medical tests, stating 'we can't track progress anymore because we're already at the ceiling.'

  3. The researchers cautioned that AI is not ready to replace physicians. While AI excels at diagnosis, it also tends to suggest unnecessary testing that could cause harm. Additionally, there is no formal framework for accountability when it comes to AI diagnoses, and the technology falls short of gaining patient trust.

Ask the AI about this article →

Get AI news like this every morning

For example, today's edition would include:

  • Meta Releases Muse Voice Transcribe, a Real-Time Speech Recognition ModelITmedia AI+ · 57m ago
  • Gartner: 55% of staff unhappy with digital upskillingITmedia AI+ · 57m ago
  • AI Users at Work 5x More Likely to Be Heavy Users, Study ShowsITmedia AI+ · 57m ago

AI-summarized, only the topics you pick — one digest a day via Email, Slack, or Discord.

Free · takes 30 seconds · unsubscribe anytimeWhat is AIToday? →

Ask AI

Ask AI anything about this article. Q&As are published on this page for other readers too.

Next articlesnc-core reduces hallucination rate by 52% on HumanEval with Qwen2.5-Coder-7B via inference-time governance layer