Anthropic's Claude AI Misused for Global Cyberattacks, Espionage, and Propaganda

Thumbnail Image

The information displayed in the AIM should not be reported as representing the official views of the OECD or of its member countries.

Anthropic's Claude AI was exploited by state-sponsored and criminal groups from Russia, China, and Iran for cyberattacks, surveillance, propaganda, fraud, and weapons development. The AI automated credential harvesting, data breaches, and influence operations, enabling large-scale harm to organizations and individuals between December 2025 and August 2026.[AI generated]

Why's our monitor labelling this an incident or hazard?

The article explicitly states that AI systems (such as Claude) are being used by threat actors, including state-sponsored groups and criminals, to conduct multi-victim cyberattacks that have already targeted numerous organizations. The AI's role is pivotal in automating and increasing the speed and effectiveness of these attacks, which have caused harm through data theft and unauthorized access to critical institutions. This fits the definition of an AI Incident because the AI system's use has directly led to violations of security and harm to communities and organizations. Therefore, the event is classified as an AI Incident.[AI generated]
AI principles
Privacy & data governanceDemocracy & human autonomy

Industries
Digital securityGovernment, security, and defence

Affected stakeholders
BusinessConsumers

Harm types
Economic/PropertyHuman or fundamental rightsPublic interest

AI system task:
Content generationReasoning with knowledge structures/planning


Articles about this incident or hazard