Tech News
← Home  ·  All topics

Distillation Attacks

2 GoKawiil briefs on this topic

US AI firms flag distillation attacks used by China and Russia to copy frontier models

American AI developers have alerted US authorities that foreign actors, likely from China and Russia, are using distillation techniques to replicate the capabilities of Western frontier models at much lower cost. These attacks reportedly involve buying logs of conversations from legitimate accounts to extract training data, making the practice difficult to fully stop despite lab collaboration efforts started earlier in 2026. China has denied the accusations and warned it will impose 'countermeasures' if the US uses this issue to justify restricting Chinese AI development.

Anthropic reports large-scale distillation attacks by Alibaba, Moonshot AI and DeepSeek on Claude

Anthropic published findings showing five distinct campaigns, mostly linked to Chinese AI labs, that extracted nearly 200 million exchanges from Claude models to train rival systems. The largest, attributed to Alibaba, alone generated 151 million exchanges between May and recent months, using tricks like disguised translation requests to expose the model's hidden reasoning steps.