Studies Warn of Security and Transparency Risks in AI Agents

Thumbnail Image

The information displayed in the AIM should not be reported as representing the official views of the OECD or of its member countries.

Multiple studies by Cambridge, MIT, and collaborators reveal that most widely used AI agents lack formal risk assessments, transparency, and adequate security measures. Only a minority disclose safety practices, raising concerns about potential vulnerabilities and uncontrolled growth that could lead to future harm if unaddressed.[AI generated]

Why's our monitor labelling this an incident or hazard?

The article clearly involves AI systems (AI agents) and discusses their development and use with insufficient safety and transparency. Although no direct harm has been reported yet, the lack of guardrails and the ability of these agents to mimic human behavior and bypass protections plausibly could lead to harms such as security breaches, misinformation, or other violations. Therefore, this is best classified as an AI Hazard, reflecting the credible risk of future AI incidents stemming from these agents' current operational state.[AI generated]
AI principles
Transparency & explainabilityRobustness & digital security

Industries
Digital security

Affected stakeholders
General public

Harm types
Human or fundamental rightsPublic interest

AI system task:
Other


Articles about this incident or hazard