By Interestana AI Editorial — AI-drafted, human-overseen. How we report
Anthropic Disrupts Chinese Labs' Claude Distillation Attacks

Anthropic announced on Thursday that it has identified and disrupted industrial-scale illicit distillation attacks targeting its Claude AI model. These sophisticated attacks were traced back to seven distinct artificial intelligence laboratories based in China. The identified entities include prominent organizations such as Alibaba, Moonshot, DeepSeek, Z.ai (also known as Zhipu), and MiniMax. The attacks involved knowledge distillation, a legitimate machine learning technique where a larger, more capable AI model acts as a 'teacher' to train a smaller model. However, in this context, the labs were allegedly using Anthropic's powerful Claude models to illicitly train their own AI systems without authorization, bypassing Anthropic's terms of service and potentially infringing on intellectual property rights.
Anthropic's security team detected these activities by monitoring for unusual patterns of access and data exfiltration that were indicative of large-scale, systematic attempts to replicate Claude's capabilities. The company stated that these distillation efforts were "industrial-scale," suggesting a significant investment of resources and computational power by the implicated labs. Such attacks aim to extract proprietary knowledge and model parameters from a teacher model to create a student model that mimics its performance, often at a fraction of the training cost and time. By engaging in this unauthorized distillation, the Chinese labs sought to gain a competitive advantage by rapidly developing advanced AI models based on Anthropic's research and development.
In response to the discovery, Anthropic took immediate action to mitigate the threat. The company stated that it has taken steps to block the identified actors and prevent further unauthorized access and misuse of its Claude models. This disruption aims to protect Anthropic's intellectual property and maintain the integrity of its AI systems. The company emphasized that knowledge distillation is a valid technique when conducted ethically and with proper licensing, but the actions of these seven labs constituted a violation of their acceptable use policies. The incident highlights the ongoing challenges in securing advanced AI models against sophisticated intellectual property theft and unauthorized replication attempts, particularly in a competitive global AI landscape.
Anthropic has not disclosed the specific versions of Claude that were targeted in these attacks, nor the exact methods used by the labs to extract the necessary data for distillation. However, the scale and systematic nature of the attacks suggest a coordinated effort. The company's proactive stance in identifying and reporting these activities underscores the increasing importance of AI security and the need for robust defenses against emerging threats. The involvement of multiple well-known Chinese AI companies also points to the intense competition and rapid development occurring within China's AI sector, where the pursuit of cutting-edge capabilities may sometimes lead to questionable practices. Anthropic's disclosure serves as a warning to other AI developers and a signal of the evolving security landscape in artificial intelligence.
Original source — read the full reporting at the publisher:
Read on The Hacker NewsGet the weekly AI digest
AI news + new model releases, weekly. Drafted by our agents, reviewed by humans.