OpenAI links China’s Moonshot AI to attempt to extract its models’ reasoning

OpenAI says it identified an attempt to extract protected reasoning from its AI models, linking part of the activity with China’s Moonshot AI.

Skip NavigationJoin ICJoin ProLivestreamMenu

  • OpenAI said the activity surged to 16,000 requests from more than 4,000 users over two days, with related activity ultimately identified across more than 15,000 users.
  • The company said operators did not breach its encryption, databases or stored user conversations.
  • The findings come weeks after Anthropic accused Chinese AI developers including Moonshot AI and Alibaba of secretly using Claude to help train their own models.

The OpenAI logo is displayed on a smartphone screen.Nurphoto | Nurphoto | Getty Images

OpenAI said it identified and disrupted a coordinated campaign in which operators attempted to extract protected reasoning from its artificial intelligence models, linking a core cluster of the activity with Chinese startup Moonshot AI, the developer of Kimi.

The activity began in early July and later surged to 16,000 requests from more than 4,000 users over two days, OpenAI said. The company ultimately identified related activity across a cluster of more than 15,000 users and said it had fully disrupted the campaign by July 28.

The findings come just weeks after OpenAI rival Anthropic accused several Chinese AI developers, including Moonshot AI and Alibaba, of secretly using its Claude model to help train their own AI systems, underscoring growing concerns among U.S. AI companies that rivals could use their models to develop competing technology more quickly and cheaply.

OpenAI described the latest activity as “adversarial distillation,” where one AI model’s outputs or reasoning are used to help train or improve another model. The company said extracting such reasoning could allow others to reproduce advanced capabilities without making the same investment in developing and safeguarding frontier models, posing potential safety and national security risks.

The operators did not breach OpenAI’s encryption, databases or stored user conversations, according to the company. Instead, they manipulated interactions with its models in an effort to reproduce hidden reasoning in a form visible to the requester.

OpenAI said it was unclear whether all the operators involved were linked to a single actor, but attributed a core cluster of the activity to individuals associated with Moonshot AI. The company said it has shared its findings with other AI developers through the Frontier Model Forum as well as government information-sharing channels.

Moonshot did not immediately respond to CNBC’s requests for comment.

Choose CNBC as your preferred source on Google and never miss a moment from the most trusted name in business news.

Leave a Reply

Your email address will not be published. Required fields are marked *

About the Author

Easy WordPress Websites Builder: Versatile Demos for Blogs, News, eCommerce and More – One-Click Import, No Coding! 1000+ Ready-made Templates for Stunning Newspaper, Magazine, Blog, and Publishing Websites.

BlockSpare — News, Magazine and Blog Addons for (Gutenberg) Block Editor

Search the Archives

Access over the years of investigative journalism and breaking reports