Get your free personalized podcast brief

We scan new podcasts and send you the top 5 insights daily.

Major AI labs are protesting that Chinese companies are "stealing" their models via distillation. However, these same labs built their foundational models by training on vast amounts of copyrighted material without permission, a practice the host calls "IP theft," undermining their public standing on the issue.

Related Insights

Proprietary labs argue against 'distillation' (using their model outputs for training) while they have built their own models on vast amounts of copyrighted data. This opposition is an anti-competitive tactic, as model outputs are not copyrightable and distillation helps smaller, open players to compete.

There is a profound hypocrisy in the AI industry's stance on intellectual property. Companies that built their foundational models by scraping the entire internet are now seeking regulatory protection to prevent others from distilling or learning from their models—mirroring how the music industry fought Napster after profiting from an open ecosystem.

The U.S. Treasury is threatening sanctions over Chinese AI labs 'distilling' U.S. models, framing a technical training process as intellectual property theft. This political reframing allows the use of powerful economic weapons outside of traditional court systems, escalating the U.S.-China AI rivalry.

US AI labs' efforts to prevent foreign rivals from distilling their models face accusations of hypocrisy. Critics point out that these labs train their own models on vast amounts of public data without permission. This "pot calling the kettle black" dynamic complicates legal and ethical arguments against industrial-scale distillation.

AI companies protest when competitors "distill" their models, calling it a violation. This stance is deeply ironic, as it mirrors the complaints of artists and creators whose work was scraped without permission to build the original models. The industry fails to acknowledge this double standard.

US officials and AI labs allege Chinese firms are engaged in industrial-scale IP theft. They reportedly use fraudulent accounts to extract capabilities from US models like Claude to train their own, creating a facade of domestic innovation.

While US AI companies navigate complex licensing deals with IP holders, Chinese firms like ByteDance appear to be using copyrighted material, such as specific actors' voices, without restriction. This lack of legal friction allows them to generate highly specific and realistic content that Western labs are hesitant to produce.

The high quality of ByteDance's C-Dance video model suggests it may be trained on copyrighted material, like David Attenborough's voice, which US labs are legally restricted from using. This freedom from IP constraints could give Chinese firms a significant competitive advantage in media generation.

Anthropic's argument that Chinese AI model distillation is 'IP theft' is a potentially fatal legal mistake. This assertion can be used against them in lawsuits from content creators like the New York Times, as Anthropic's own models are built by 'distilling' public content, effectively confessing their product is based on stolen IP.

The US accuses China of "distillation"—querying American AI models millions of times to reverse-engineer their logic and capabilities. This marks a shift from commercial competition to industrial-scale intellectual property theft, escalating the geopolitical conflict beyond government rhetoric.

AI Labs Built on IP Theft Now Condemn Chinese Rivals for Using Similar Methods | RiffOn