6 Sources
[1]
Penguin Random House is adding an AI warning to its books' copyright pages
Penguin Random House, the world's largest trade publisher, will be adding language to the copyright pages of its books to prohibit the use of those books to train AI. The Bookseller reports that new books and reprints of older titles from the publisher will now include the statement, "No part of
[2]
Penguin Random House copyright pages will now forbid AI training
Credit: Idrees Abbas/SOPA Images/LightRocket via Getty Images Penguin Random House (PRH), the largest of the Big Five publishing imprints, is pushing back against its published works being used to train AI. As first reported by The Bookseller, PRH has changed its copyright wording to target AI.
[3]
Major Publisher Penguin Random House Blocks AI Training on Its Books
Penguin Random House, one of the world's largest publishers, has taken action to block firms from training AI systems on its huge portfolio, publishing trade The Bookseller reports. AI firms often trawl or "scrape" sources like fiction and non-fiction books, newspapers, and social media to train
[4]
Penguin Random House amends its copyright rules to protect authors from AI
The new warning prohibits AI companies from training on the publisher's books. Artificial intelligence makers have faced a mountain of criticism for borrowing from the work of others to train its models. Now the world's largest publishing house is taking steps to ensure its authors don't have
[5]
Penguin Adds a Do-Not-Scrape-for-AI Page to Its Books
The world's largest publishing house doesn't want its copyrighted work used to train AI models. Taking a firm stance against tech companies' unlicensed use of its authors' works, the publishing giant Penguin Random House will change the language on all of its books' copyright pages to expressly
[6]
Penguin Random House books now explicitly say 'no' to AI training
What gets printed on that page might be a warning shot, but it also has little to do with actual copyright law. The amended page is sort of like Penguin Random House's version of a robots.txt file, which websites will sometimes use to ask AI companies and others not to scrape their content. But
Share
Copy Link
Penguin Random House, the world's largest trade publisher, has updated its copyright pages to prohibit the use of its books for training AI systems, marking a significant move in the ongoing debate over AI and copyright.

Penguin Random House (PRH), the world's largest trade publisher, has taken a significant step to protect its authors' intellectual property by adding new language to the copyright pages of its books. The updated text explicitly prohibits the use of their publications for training artificial intelligence technologies or systems
1
.The amended copyright statement now reads: "No part of this book may be used or reproduced in any manner for the purpose of training artificial intelligence technologies or systems"
2
. This change will be implemented across all PRH imprints, affecting both new titles and reprints of older works.While this update doesn't necessarily alter the legal status of the texts, it represents a clear stance against the unauthorized use of copyrighted material in AI development. PRH is the first among the "Big Five" English language publishers to take such a public action
3
.The move aligns with a European Parliament directive that grants copyright holders the right to protect their material from text or data mining by AI firms, provided they opt out of such usage
3
. This could have far-reaching implications for AI companies that rely on vast amounts of text data for training their models.Copyright lawyer Chien-Wei Lui supports PRH's decision, stating that it helps preserve the value of author content and encourages proper licensing practices
3
. However, not all publishers are taking the same approach. Some, like Wiley, Oxford University Press, and Taylor & Francis, have signed agreements allowing their content to be used for AI training under certain conditions3
.This development is part of a larger trend of content creators and publishers pushing back against the use of their work in AI training. In late 2023, The New York Times sued OpenAI and Microsoft for copyright infringement, claiming that millions of its articles were used to train AI models without permission
4
.Related Stories
PRH's stance could significantly impact the AI industry. Books are considered high-quality training data due to their well-written and fact-checked content. If other major publishers follow suit, AI companies may be forced to either pay for access to quality content or rely on potentially lower-quality, freely available internet data
5
.Matthew Sag, an AI and copyright expert, suggests that this move by PRH could lead to a new model where publishers and websites monetize access to their content for AI training purposes. This compromise could allow AI companies to continue training on the "open Internet" while enabling content owners to benefit from the use of their work in AI development
5
.Summarized by
Navi
19 Nov 2024•Business and Economy

31 Mar 2026•Policy and Regulation

29 Jun 2025•Technology
