Facebook
Britain's News Portal
Around The Clock
BREAKING
Loading latest headlines…

Anthropic Alleges Escalation in AI Model Distillation Attacks

Anthropic has released a report alleging persistent and escalating distillation attacks by China-based AI companies, observing nearly 200 million exchanges linked to these efforts.

  • Anthropic's report details five separate campaigns, attributing the bulk of activity to Alibaba.
  • The attacks aim to extract a model's 'chain of thought' to train smaller models on reasoning ability.
  • One campaign attributed to Moonshot AI reportedly routed requests from the Chinese military.

Anthropic has released a new report alleging persistent distillation attacks by China-based AI companies, which have reportedly escalated in recent months. The company observed nearly 200 million exchanges linked to these attacks, attributed to five separate campaigns.

The report states that unauthorised labs have developed increasingly sophisticated methods to circumvent defenses and harvest capabilities from US frontier models. These campaigns reportedly targeted Claude's capabilities, including agentic capabilities, tool use, coding, data analysis, and logical reasoning.

The bulk of distillation attempts, 151 million exchanges between May and July 2026, were attributed to Alibaba. Anthropic described this as the largest wholesale distillation effort it has observed, peaking at nearly three million exchanges per day. Another campaign from Moonshot AI reportedly routed requests directly from the Chinese military.

Related Articles

Get the news that matters.

Join thousands of readers getting the best of British news straight to their inbox.