Anthropic Raises Alarm Over Distillation Attacks from China-Based AI Labs

Anthropic reveals a surge in distillation attacks on its AI models, targeting valuable capabilities and potentially aiding rival technologies.

AI2Day Newsdesk2 min read
Full-frame edge-to-edge photoreal news-editorial image of a modern laptop screen showing an abstract browser window with a glowing extension icon in the toolbar
Share

Key points

  • Anthropic reported nearly 200 million distillation attacks by China-based AI firms as of October 2023.
  • Alibaba led the attacks with 151 million exchanges between May and July 2026.
  • Moonshot AI's campaign allegedly involved requests from the Chinese military.

What happened?

Anthropic, an American AI company, has reported extensive distillation attacks from China-based AI companies. These attacks, which have intensified over several months, involve methods to bypass Anthropic's defenses and extract critical capabilities from their models. A distillation attack is a technique where attackers trick a model into revealing its

© 2026 AI2Day