AI News Feed
Market watch
Companies

OpenAI says it disrupted campaign to extract model reasoning, links core activity to Moonshot AI

OpenAI says it disrupted a campaign to extract protected AI reasoning and links a core cluster to China's Moonshot AI.

The activity began in early July and later surged to 16,000 requests from more than 4,000 users over two days, OpenAI said. The company said it ultimately identified related activity across a cluster of more than 15,000 users and had fully disrupted the campaign by July 28.

OpenAI described the activity as "adversarial distillation," in which one AI model's outputs or reasoning are used to help train or improve another model. The company said extracting such reasoning could allow others to reproduce advanced capabilities without making the same investment in developing and safeguarding frontier models, posing potential safety and national security risks.

According to OpenAI, the operators did not breach its encryption, databases, or stored user conversations. Instead, they manipulated interactions with its models in an effort to reproduce hidden reasoning in a form visible to the requester.

OpenAI said it was unclear whether all the operators involved were linked to a single actor, but it attributed a core cluster of the activity to individuals associated with Moonshot AI. The company said it shared its findings with other AI developers through the Frontier Model Forum and government information-sharing channels.

Moonshot did not immediately respond to CNBC's requests for comment.

The findings come just weeks after OpenAI rival Anthropic accused several Chinese AI developers, including Moonshot AI and Alibaba, of secretly using its Claude model to help train their own AI systems. The accusation underscored growing concerns among U.S. AI companies that rivals could use their models to develop competing technology more quickly and cheaply.