
The information displayed in the AIM should not be reported as representing the official views of the OECD or of its member countries.
A UK fintech firm, Saturn, found that major AI chatbots—including ChatGPT, Gemini, Claude, and Copilot—gave incorrect or incomplete financial advice in 57% of over 10,000 tested answers. Error rates rose to 88% for complex queries, exposing users to potential financial harm due to misinformation.[AI generated]
Why's our monitor labelling this an incident or hazard?
The article clearly involves AI systems (chatbots like ChatGPT, Gemini, Claude) used for financial advice. The study shows these AI systems often give wrong answers, which can mislead users and cause financial harm. This meets the definition of an AI Incident because the AI's use has directly or indirectly led to harm (financial loss) to people. The harm is realized, not just potential, as the article cites specific examples of incorrect advice that could cause severe financial outcomes. Therefore, this is an AI Incident.[AI generated]