Study Finds ChatGPT Health AI Fails in Emergency Medical Triage

Thumbnail Image

The information displayed in the AIM should not be reported as representing the official views of the OECD or of its member countries.

An independent study found that ChatGPT Health, an AI medical guidance tool used by millions, failed to recommend emergency care in over half of serious cases and inconsistently flagged suicide risks. Researchers warn these triage failures pose significant health risks for users relying on the AI for urgent decisions.[AI generated]

Why's our monitor labelling this an incident or hazard?

ChatGPT Health is an AI system providing health guidance. The study shows that its use has led to undertriage of serious medical emergencies and inconsistent suicide-crisis alerts, which can cause harm to users' health by delaying or preventing necessary emergency care. Although the harm is indirect (due to the AI's failure to recommend appropriate emergency responses), the risk and actual instances of undertriage constitute injury or harm to persons. Therefore, this qualifies as an AI Incident under the framework, as the AI system's use has directly or indirectly led to harm to health.[AI generated]
AI principles
SafetyAccountability

Industries
Healthcare, drugs, and biotechnology

Affected stakeholders
Consumers

Harm types
Physical (injury)Physical (death)

AI system task:
Interaction support/chatbotsOrganisation/recommenders


Articles about this incident or hazard