U.S. cybersecurity and intelligence agencies accuse China-based AI companies of carrying out “systematic extraction” of proprietary capabilities from leading American frontier models using distillation techniques. The agencies describe the activity as occurring at an “industrial-scale” and as a central element of the companies’ AI development strategies rather than a minor or supplementary practice.
The agencies—reported as including the NSA, CISA, and the FBI—say the firms route and process large volumes of requests through multiple accounts, APIs, proxies, cloud providers, and third-party aggregators. According to the filings described by outlets, the goal is to replicate or learn capabilities from models such as ChatGPT, Claude, Gemini, and Grok, based on observations spanning from at least 2024.
Outlets differ mainly in the level of detail they provide about specific companies and alleged methods. Slashdot, for example, describes how DeepSeek and Moonshot AI allegedly used distillation to produce synthetic training data and new model versions, including work tied to capabilities like coding, agentic behavior, and question-answer optimization. The Hacker News frames the broader allegation in terms of extraction of proprietary functionalities and capabilities.