Skip to main content
AI

OpenAI Model Outperforms Doctors in Emergency Room Diagnosis Trial

OpenAI Model Outperforms Doctors in Emergency Room Diagnosis Trial Image: primary
A Harvard study published in the journal Science found that an artificial intelligence system outperformed human doctors in emergency room triage diagnosis accuracy. In one experiment involving 76 patients at a Boston hospital, OpenAI's o1 reasoning model correctly identified the exact or very close diagnosis in 67 percent of cases, compared with 50 to 55 percent for teams of human doctors. The AI had access to the same electronic health records that doctors typically review, including vital sign data, demographic information, and brief notes from nurses about why the patient arrived. When more detailed information was available, the AI's accuracy rose to 82 percent, compared with 70 to 79 percent for specialist doctors. That gap was not statistically significant, suggesting the technology excels most where decisions must be made rapidly with minimal information. The AI also outperformed human doctors on longer-term treatment planning. When asked to examine five clinical case studies and develop treatment plans, the model scored 89 percent compared with 34 percent for doctors using conventional resources such as search engines. In one notable case, a patient presented with a blood clot to the lungs and worsening symptoms. Human doctors suspected the blood thinners were failing, but the AI identified lupus as the underlying cause of lung inflammation, a conclusion that was subsequently confirmed. The researchers cautioned that the AI was tested only on text-based information and was not evaluated on visual cues such as a patient's appearance or level of distress. Lead author Arjun Manrai of Harvard Medical School said the findings do not mean AI will replace doctors, but that the technology represents a profound change that will reshape medicine. Co-lead author Adam Rodman said the model adds an artificial intelligence component to a triadic care system involving doctor, patient, and AI. Approximately one in five US physicians already use AI to assist with diagnosis, according to recent research. In the United Kingdom, 16 percent of doctors use AI daily and 15 percent use it weekly, with clinical decision-making cited as a common application in a Royal College of Physicians survey. British doctors' primary concerns are AI error and liability risks, particularly given the absence of a formal accountability framework.
Sources
In this story
Published by Tech & Business, a media brand covering technology and business. This story was sourced from The Guardian and reviewed by the T&B editorial agent team.
Back to Newswire
Keep reading
Full wire
AI Products
AI Products

AWS makes GPT-6 Astra generally available on Bedrock

AWS said OpenAI's GPT-6 Astra is now generally available through Amazon Bedrock. The company says customers can call the model through supported Bedrock APIs or configure ChatGPT Work and Codex to use it on Bedrock. AWS describes ...

AI Products
AI Products

OpenAI reportedly buys Glass Imaging for more than $300 million

OpenAI has bought smartphone-camera company Glass Imaging in a deal worth more than $300 million, The Wall Street Journal reported, according to TechCrunch. Glass Imaging uses neural networks trained on individual smartphone camer...

AI Products
AI Products

xAI and X resolve Apple antitrust lawsuit

xAI and X Corp. said they resolved their antitrust lawsuit against Apple, which had accused the iPhone maker of favoring OpenAI's ChatGPT over other chatbot makers. The supplied summary does not disclose the resolution's terms or ...

AI Security
AI Security

OpenAI confirms agents were involved in May RubyGems incident

OpenAI confirmed that its agents used RubyGems during a May incident in which more than 2,000 packages were submitted over two days, according to reporting based on public package records and researcher analysis. The packages used...

AI Products
AI Products

Large AI customers curb Fable use over retention concerns

Some large AI customers are restricting which models employees may use or which tasks they may use them for over fears proprietary information could be used in training, The Information reported. Nvidia uses Fable only for work th...

AI
AI

Apple code points to broader Siri model interoperability

MacRumors reported that private frameworks in iOS 27 and macOS Golden Gate show Apple has built Siri mechanisms for third-party model delegation and inference providers. A demonstration showed a Claude extension handling a reques...