OpenAI links actors tied to Moonshot AI to reasoning-extraction campaign
OpenAI says a coordinated campaign attempted to extract its models' hidden 'chain of thought' reasoning in July, with actors associated with China-based Moonshot AI at the center. Activity started July 1 and spiked to 16,000 extraction-pattern requests from over 4,000 users on July 24-25, before OpenAI says it fully disrupted the campaign by July 28. OpenAI says the encryption protecting this reasoning data was not broken and no user conversation database was compromised.
GoKawiil's interpretation of the reporting above, not reported fact.
The incident highlights growing concern among AI labs about rivals using 'adversarial distillation' to cheaply replicate advanced reasoning capabilities by harvesting outputs rather than building models from scratch. OpenAI's disclosure, without specifying success rates or confirming Moonshot AI's direct involvement, suggests the company wants to publicize the threat while leaving some details deliberately vague. The response—patching a replay vulnerability and adding cross-tenant protections—indicates OpenAI sees protecting proprietary reasoning as a competitive and security priority.
- OpenAI reports a July campaign to extract protected 'chain of thought' reasoning from its models, linking it to actors associated with Moonshot AI.
- Over 4,000 users generated 16,000 extraction-pattern requests during spikes on July 24-25, but OpenAI says the encryption itself was not broken.
- OpenAI says it disrupted the campaign by July 28 and has since patched a replay vulnerability and added stronger cross-tenant reasoning protections.
Source: tomshardware.com — Shane Downing, 2026-10-01
Published there as: “OpenAI says actors linked to China-based Moonshot AI spearheaded a campaign to extract its models’ hidden reasoning”
Read the original report → The summary and analysis above are GoKawiil's own, written from reporting by the source above. Facts and quotes belong to the original publisher.