3 Sources
[1]
Reddit keeps weird DMCA lawsuit against web scraper alive despite Google's loss
On Friday, a judge largely denied a motion to dismiss from a web scraper, SerpApi, which is accused of conspiring with Perplexity AI to illegally scrape copyrighted Reddit content from Google search results. In his opinion, US District Judge Paul A. Engelmayer said that at this early stage, Reddit
[2]
Reddit's AI copyright lawsuit against Perplexity can move forward.
A judge rejected Perplexity's motion to dismiss the lawsuit, which accuses the AI startup and three data-scraping services of vacuuming up Reddit's content without permission. Ben Lee, Reddit's chief legal officer, says in a statement to The Verge that the ruling "brings us one step closer to
[3]
Perplexity AI loses bid to toss Reddit lawsuit over data scraping
July 31 (Reuters) - A Manhattan federal judge on Friday rejected most of Perplexity AI's bid to dismiss a lawsuit brought by Reddit (RDDT.N), opens new tab accusing it of violating U.S. copyright law by scraping the online discussion platform's data to train Perplexity's AI-powered search
Share
Copy Link
A Manhattan federal judge rejected Perplexity AI's motion to dismiss Reddit's DMCA lawsuit over data scraping. Reddit accuses Perplexity and third-party scrapers like SerpApi of bypassing protections to illegally scrape copyrighted Reddit content from Google search results for AI training data.
A Manhattan federal judge has rejected most of Perplexity AI's bid to dismiss a Reddit lawsuit accusing the AI startup of violating copyright law through unauthorized scraping of Reddit's content
3
. US District Judge Paul A. Engelmayer ruled on Friday that Reddit has plausibly demonstrated that Perplexity AI and data scraping services conspired to illegally scrape copyrighted Reddit content from Google search results1
. The decision allows Reddit to continue pressing claims that Perplexity and three data scrapers—SerpApi, Oxylabs, and AWMProxy—unlawfully circumvented protective measures to access content for AI training data3
.
Source: The Verge
The ruling comes less than two weeks after another court dismissed a similar action raised by Google, finding insufficient proof that rights holders like Reddit had authorized the search engine to prevent scraping
1
. Judge Engelmayer distinguished Reddit's case by noting the company went beyond bare allegations, arguing its licensing agreement with Google directly prohibits certain uses of Reddit data now being accessed through anti-circumvention measures employed by third-party scrapers1
. Reddit successfully argued that when it licenses content to partners like Google, those partners agree to delete posts flagged when users remove content—a promise that millions of monthly post deletions depend upon1
.Judge Engelmayer ruled that Reddit has standing to sue for Perplexity's alleged misuse of its users' content, despite the platform not owning the copyrighted material itself
3
. This addresses a key question raised by DMCA experts about whether Reddit, as neither the copyright owner nor the creator of Google's protective technology, could bring such claims1
. The decision suggests courts may accept that platforms can enforce protections on behalf of their user communities when licensing agreements are in place.
Source: Ars Technica
Reddit has asked the court for multiple forms of relief: an injunction blocking SerpApi and Perplexity AI access to both Reddit and Google websites, another injunction stopping circumvention of Google SearchGuard, and a third preventing the defendants from using previously scraped data
1
. The platform also seeks unspecified monetary damages3
. Ben Lee, Reddit's chief legal officer, stated the ruling brings the company closer to holding bad actors accountable, emphasizing that Reddit supports responsible access to public content but opposes companies that bypass protections and profit off communities without permission2
.Perplexity AI responded defiantly, stating that Reddit's suit claims the right to control access to public web pages it doesn't own, using a security tool it didn't build, on behalf of users it hasn't asked
3
. The company vowed to defend the open internet and predicted victory. SerpApi similarly argued that Google and Reddit are trying to use the DMCA to wall off the open Internet by retroactively claiming control over content they didn't author and don't own1
. SerpApi's attorney stated the company accesses public search results, not Reddit's platform, and that public information doesn't become protected simply because a platform wants to charge for it3
.Related Stories
If Reddit wins this AI copyright lawsuit, the platform could be better positioned to force all AI scrapers to enter licensing agreements
1
. Reddit, which features thousands of interest-based subreddit communities, has already licensed its content to Google, OpenAI, and others for AI training3
. The lawsuit claims Reddit is the most commonly cited source for AI-generated answers to user questions, making control over this data particularly valuable3
. This case joins many filed by content owners including authors, music labels, and news outlets against tech companies over alleged misuse of copyrighted material for AI systems3
.Judge Engelmayer noted that SerpApi and Perplexity AI may still prove through discovery that Reddit never authorized Google to protect its content in Google search results
1
. SerpApi may also strengthen its defense by proving that publicly accessible content in search results isn't protected by the Copyright Act1
. The fight appears far from over, with fundamental questions about who controls public web content and how anti-circumvention measures apply to AI training still unresolved. Reddit has also filed a separate data-scraping lawsuit against Anthropic that remains ongoing in California state court3
.Summarized by
Navi
22 Oct 2025•Policy and Regulation

05 Jun 2025•Policy and Regulation

01 Aug 2024

1
Policy and Regulation

2
Technology

3
Technology
