Meta's AI Moderation Fails to Curb Anti-Trans Hate Speech, GLAAD Reports

Thumbnail Image

The information displayed in the AIM should not be reported as representing the official views of the OECD or of its member countries.

GLAAD reports that Meta's AI-driven content moderation systems are failing to remove extreme anti-trans hate speech across Facebook, Instagram, and Threads. Despite policies banning such content, harmful posts—including slurs and calls for violence—remain widespread, contributing to real-world harm against LGBTQ+ communities.[AI generated]

Why's our monitor labelling this an incident or hazard?

Meta's content moderation relies heavily on AI systems to detect and remove harmful content. The report highlights that these AI systems are allowing anti-trans hate speech and calls for violence to remain online, which has resulted in documented real-world harms to LGBTQ+ people. This constitutes an AI Incident because the AI system's malfunction or inadequacy in moderating content has directly or indirectly led to harm to communities and violations of rights.[AI generated]
AI principles
SafetyFairnessRespect of human rights

Industries
Media, social platforms, and marketing

Affected stakeholders
Other

Harm types
PsychologicalHuman or fundamental rights

Business function:
Monitoring and quality control

AI system task:
Recognition/object detection


Articles about this incident or hazard