AI Is Already Beating Human Doctors in Medical Tests (Elizabeth Nolan Brown, August/September 2026, reason)
In a recent series of experiments led by researchers at the Beth Israel Deaconess Medical Center and Harvard Medical School, a preview of the OpenAI large language model (LLM) known as o1 bested human physicians in multiple tests of clinical and diagnostic reasoning.
“We tested the AI model against virtually every benchmark, and it eclipsed both prior models and our physician baselines,” said Harvard Medical School professor Arjun K. Manrai, one of the study’s senior authors.
