New Zealand Develops AI Tool to Redirect Extremist Users to Deradicalization Support

Thumbnail Image

The information displayed in the AIM should not be reported as representing the official views of the OECD or of its member countries.

ThroughLine, contracted by OpenAI, Anthropic, and Google, is developing an AI system in New Zealand to detect users exhibiting violent extremist tendencies on platforms like ChatGPT and redirect them to human and chatbot-based deradicalization support. The tool aims to prevent harm but is still in testing.[AI generated]

Why's our monitor labelling this an incident or hazard?

The event involves the use of an AI system (a chatbot and detection system) to identify and intervene with users showing violent extremist tendencies, which is a clear AI system involvement. However, the article does not report any actual harm or incident caused by the AI system; rather, it discusses the development and testing of a tool aimed at preventing harm. Therefore, it fits the definition of an AI Hazard, as the AI system's use could plausibly lead to preventing or managing harm related to extremism, but no incident has yet occurred. The article also discusses the broader context of safety concerns and potential misuse but does not describe a realized AI Incident or complementary information focused on responses to a past incident.[AI generated]
AI principles
Privacy & data governanceRespect of human rights

Industries
Government, security, and defence

Business function:
Citizen/customer service

AI system task:
Event/anomaly detectionInteraction support/chatbots


Articles about this incident or hazard