OpenAI Thwarts Large-Scale AI Model Cloning Attempt Linked to Moonshot AI

Thumbnail Image

The information displayed in the AIM should not be reported as representing the official views of the OECD or of its member countries.

OpenAI detected and stopped a coordinated campaign involving over 15,000 users attempting to clone ChatGPT's capabilities using distillation techniques. The operation, partly attributed to China's Moonshot AI (developer of Kimi), aimed to extract and replicate proprietary model behaviors, raising concerns over intellectual property violations and security risks.[AI generated]

Why's our monitor labelling this an incident or hazard?

The event involves the use and misuse of AI systems (ChatGPT and related models) in a coordinated attempt to clone their capabilities through adversarial distillation. This activity exploited vulnerabilities and involved a large number of users, leading to direct risks to safety and national security, which are harms under the AI Incident criteria. OpenAI's intervention stopped the harm, but the incident itself involved realized harm and security risks. Therefore, this qualifies as an AI Incident rather than a hazard or complementary information.[AI generated]
AI principles
AccountabilityRobustness & digital security

Industries
Digital security

Affected stakeholders
Business

Harm types
Economic/Property

Business function:
Research and development

AI system task:
Content generation


Articles about this incident or hazard