ChatGPT Health Delivers Inaccurate and Alarmist Health Assessments from Apple Health Data

Thumbnail Image

The information displayed in the AIM should not be reported as representing the official views of the OECD or of its member countries.

OpenAI's ChatGPT Health, designed to analyze personal health data from apps like Apple Health, has produced inconsistent and alarmist health assessments, as revealed by tests from The Washington Post. Medical experts found these AI-generated diagnoses to be unfounded, raising concerns about user confusion, anxiety, and potential health risks from misleading outputs.[AI generated]

Why's our monitor labelling this an incident or hazard?

The article explicitly involves an AI system (ChatGPT Salud) processing personal health data to generate health assessments. The AI system's malfunction—providing inaccurate and inconsistent diagnoses—directly risks harm to users' health by potentially misleading them about their medical condition. The harm is related to injury or harm to health (a), as users might rely on incorrect AI-generated diagnoses. The event is not merely a potential risk but documents actual erroneous outputs and misleading conclusions, fulfilling the criteria for an AI Incident rather than an AI Hazard or Complementary Information. The lack of regulation and the possibility of user reliance on these faulty outputs further supports the classification as an AI Incident.[AI generated]
AI principles
SafetyTransparency & explainability

Industries
Healthcare, drugs, and biotechnology

Affected stakeholders
Consumers

Harm types
PsychologicalPhysical (injury)

Business function:
Citizen/customer service

AI system task:
Forecasting/prediction


Articles about this incident or hazard