Urgent.News

What's breaking now, across thousands of outlets.

AI

US claims Chinese AI companies’ core AI strategy is distilling American models

Spooks and CISA point to ‘industrial-scale distillation’ by DeepSeek, Alibaba, and other Chinese players

US claims Chinese AI companies’ core AI strategy is distilling American models

Two US intelligence agencies, the NSA, FBI, and CISA, have accused Chinese AI companies of conducting widespread and malicious "distillation activities" aimed at extracting valuable functionalities from American AI models. The agencies claim that this process is at the core of China's AI development strategy. Distillation involves using a smaller model to learn from a larger, more complex model, allowing the smaller model to improve its performance over time.

However, commercial providers often forbid such activity due to the substantial investment in creating the models. Chinese AI companies allegedly bypass U.S. terms of use by routing requests through multiple pathways, including APIs, cloud providers, and third-party aggregators. They also use "transfer stations" to circumvent geographic restrictions and evade safeguards.

The advisory cites instances of Chinese companies distilling models to create synthetic data, falsely claiming to have achieved high-quality results with minimal resources. This could have severe implications for American AI companies that have invested heavily in infrastructure. The agencies recommend that AI companies improve their detection and defense mechanisms against distillation attacks, such as subtly altering responses for suspected malicious activity and correlating activity across different providers and platforms.

Written by urgent.news from The Register's reporting — not their text. Machine-written — may contain errors; check the original before relying on it.

This story

This is one outlet's version. Read the fullest account.

Read the original at theregister.com →

More in AI

Translating 300-Page Books with Claude: Taming Token Limits and Context Windows

How we chunk long-form content and maintain translation quality with Claude API When we launched LectuLibre, our AI-powered book translation platform, we thought the hard part would be fine-tuning…

  • Claude's API has 200,000-token context window and 4,096-token output limit.
  • Chunking system splits 300-page books into manageable pieces fitting token limits.
  • Running context with previous translated segment and glossary improves translation consistency.

More from Wednesday 9 September →