Claude's Capabilities Extracted at Large Scale? Anthropic Accuses Ali-Related Parties of 'Distilling' Models
Anthropic alleges that entities associated with Alibaba and its AI lab Qwen used nearly 25,000 fraudulent accounts to extract capabilities from its Claude AI model in what it calls the "largest known" model distillation attack. The alleged incident occurred between April 22 and June 5, involving over 28.8 million interactions with Claude. Model distillation involves using a powerful model's outputs to train another model, potentially replicating abilities like software engineering and agent reasoning without stealing the underlying code.
The accusations emerge amid heightened US AI export controls and the Pentagon's listing of Alibaba as a "Chinese military company." Anthropic detailed the claims in a June 10 letter to the US Senate Banking Committee, urging better threat intelligence sharing. While not conclusively proving direct Alibaba involvement or successful capability replication, the case highlights growing concerns over AI model outputs as contested assets. The incident may push for stricter controls on model access and user verification, increasing compliance costs for AI firms and potentially limiting Chinese companies' access to advanced foreign models.
marsbit06/25 06:29