2 Sources
[1]
Meta defends its vast book torrenting: We're just a leech, no proof of seeding
Just because Meta admitted to torrenting a dataset of pirated books for AI training purposes, that doesn't necessarily mean that Meta seeded the file after downloading it, the social media company claimed in a court filing this week. Evidence instead shows that Meta "took precautions not to 'seed'
[2]
Meta defends using pirated material, claims it's legal if you don't seed content
Configuration settings were modified "so that the smallest amount of seeding possible could occur". Meta claimed in a court filing this week that despite torrenting an 82 TB dataset of pirated, copyrighted material from shadow libraries to train its LLaMA AI models, that employees "took
Share
Copy Link
Meta claims it didn't seed pirated books used for AI training, sparking debate on copyright infringement and data acquisition methods in AI development.

Meta, the social media giant, is embroiled in a legal battle over its use of pirated books to train its AI models. In a recent court filing, Meta defended its actions by claiming that while it did torrent a dataset of pirated books, it took precautions not to "seed" any downloaded files
1
.Meta admitted to torrenting an 82 TB dataset of pirated, copyrighted material from shadow libraries to train its LLaMA AI models. However, the company insists that there is no evidence of "seeding" - the act of sharing a torrented file after the download completes
2
.The lawsuit, filed by authors including Richard Kadrey, Sarah Silverman, and Ta-Nehisi Coates, alleges that Meta unlawfully copied and distributed their works through AI outputs. Meta's defense hinges on the lack of evidence of seeding, arguing that downloading copyrighted content isn't illegal, but distribution is
1
.Despite Meta's claims, there is testimony that might challenge their defense:
2
.1
.Meta is attempting to dismiss the authors' claim under California's Computer Data Access and Fraud Act (CDAFA), arguing it's preempted by copyright law. The authors contend that Meta's "decision to bypass lawful acquisition methods" constitutes a separate CDAFA violation
1
.Related Stories
This case highlights the ongoing tension between AI development and copyright law. Similar lawsuits have been filed against other AI companies, including OpenAI and Microsoft, over the use of copyrighted material for training large language models
2
.The outcome of this case could have far-reaching implications for the AI industry, potentially setting precedents for how companies can legally acquire and use data for AI training. It also raises questions about the ethics of using pirated material for technological advancement
1
2
.As the court battle continues, no final decision has been made. Meta is expected to fight the seeding claims at summary judgment, and any decision is likely to face appeals, suggesting a long legal process ahead
1
2
.Summarized by
Navi
11 Mar 2025•Policy and Regulation

10 Mar 2026•Policy and Regulation

08 Feb 2025•Technology
1
Science and Research

2
Technology
3
Technology