6 Sources
[1]
Meta mocked for raising "Bob Dylan defense" of torrenting in AI copyright fight
Authors think that Meta's admitted torrenting of a pirated books data set used to train its AI models is evidence enough to win their copyright fight -- which previously hinged on a court ruling that AI training on copyrighted works isn't fair use. Moving for summary judgment on a direct copyright
[2]
Judge Allows Authors' AI Copyright Case Against Meta to Proceed
Meta's AI training practices are about to face legal scrutiny, as a judge has allowed a copyright infringement case against the company to proceed. The lawsuit, filed by authors Richard Kadrey and Christopher Golden and comedian Sarah Silverman in July 2023, accuses Meta of using material from
[3]
Meta may have illegally removed copyright info in AI corpus
Facebook giant allegedly didn't want neural networks to emit results that would give the game away A judge has found Meta must answer a claim it allegedly removed so-called copyright management information from material used to train its AI models. The Friday ruling by Judge Vince Chhabria
[4]
Piracy lawsuit against Meta could set precedent for torrenting copyrighted works in AI training
A hot potato: Meta is embroiled in a ground-breaking AI lawsuit that could change how courts view copyright law. The case seems open-and-shut from the plaintiffs' view. However, if a judge sees otherwise, it could set a monumental precedent allowing corporations to pirate copyrighted material to
[5]
Did Meta used torrented books to train AI?
Meta is being accused of violating copyright laws through its admitted torrenting of a pirated books dataset utilized for training its AI models, according to a summary judgment filing in a US district court in California reported by ArsTechnica. In their legal filing, the authors argued that
[6]
Meta vs. Kadrey Sets Precedent for AI Copyright Battles
Judge Vince Chhabria determined that Kadrey et al's Digital Millennium Copyright Act (DMCA) claim could move ahead. Credit: Pexels. A California Judge has partially granted Meta's motion to dismiss a lawsuit brought by a group of authors, including novelist Richard Kadrey and comedian Sarah
Share
Copy Link
Meta is embroiled in a lawsuit accusing the company of using torrented copyrighted books to train its AI models, potentially setting a precedent for how courts view copyright law in AI development.

Meta, the parent company of Facebook and Instagram, is facing a significant legal challenge over its artificial intelligence (AI) training practices. A group of authors, including Richard Kadrey, Sarah Silverman, and Christopher Golden, have filed a lawsuit accusing Meta of copyright infringement by using their copyrighted works to train its Llama AI model without permission
2
.The plaintiffs allege that Meta resorted to torrenting terabytes of pirated book data to train its AI models after attempts to download pirated books individually proved too slow and strained Meta's networks
1
. This decision, they argue, was made with full awareness of the legal risks involved, as torrenting has long been associated with copyright infringement4
.The authors claim that Meta's actions constitute clear copyright infringement, stating, "Whatever the merits of generative artificial intelligence, or GenAI, stealing copyrighted works off the Internet for one's own benefit has always been unlawful"
1
. They further allege that Meta attempted to conceal its torrenting activities by using Amazon Web Services rather than its own infrastructure4
.The plaintiffs mockingly refer to Meta's stance as the "Bob Dylan defense," citing lyrics from Dylan's "Sweetheart Like You": "Steal a little and they throw you in jail / Steal a lot and they make you king"
1
5
. They argue that Meta's use of peer-to-peer (P2P) file sharing to obtain copyrighted material cannot be considered fair use1
.Meta has admitted to using the Books3 dataset, which contains 195,000 copyrighted books, to train its Llama 1 large language model
3
. However, the company maintains that its actions fall under fair use doctrine4
.The outcome of this case could have far-reaching implications for AI development and copyright law. If the court rules in favor of Meta, it could set a precedent allowing AI developers to use copyrighted material for training without compensation to intellectual property owners
4
. Conversely, a ruling in favor of the authors could strengthen similar cases and potentially lead to copyright reform4
.Related Stories
Judge Vince Chhabria, who is presiding over the case, has allowed it to proceed, stating that "Copyright infringement is obviously a concrete injury sufficient for standing"
2
. However, he has also expressed some unfamiliarity with torrenting terminology, which may influence how the case proceeds4
.The authors have filed for a partial summary judgment, arguing that Meta's use of torrenting leaves no room for legal ambiguity
4
. Judge Chhabria is scheduled to evaluate these claims in a hearing on May 1, considering whether ruling on the summary judgment at this stage might be unfair to Meta1
.This case is part of a larger trend of legal challenges facing AI companies over copyright issues. The New York Times has sued OpenAI and Microsoft, News Corp. has sued Perplexity, and several Canadian news organizations have sued OpenAI
2
. A recent ruling in favor of Thomson Reuters against Ross Intelligence has already suggested that indiscriminate ingestion of copyrighted material for AI training may have financial consequences3
.As the legal battle unfolds, the tech industry and copyright holders alike are closely watching this case, which could significantly shape the future landscape of AI development and intellectual property rights in the digital age.
Summarized by
Navi
[3]
[4]
[5]
1
Science and Research

2
Technology
3
Technology