Reddit Lawsuit Against Perplexity AI and SerpApi Survives Dismissal in Data Scraping Battle

Reviewed byNidhi Govil

3 Sources

Share

A Manhattan federal judge rejected Perplexity AI's motion to dismiss Reddit's DMCA lawsuit over data scraping. Reddit accuses Perplexity and third-party scrapers like SerpApi of bypassing protections to illegally scrape copyrighted Reddit content from Google search results for AI training data.

Reddit Scores Early Win in AI Copyright Lawsuit

A Manhattan federal judge has rejected most of Perplexity AI's bid to dismiss a Reddit lawsuit accusing the AI startup of violating copyright law through unauthorized scraping of Reddit's content

3

. US District Judge Paul A. Engelmayer ruled on Friday that Reddit has plausibly demonstrated that Perplexity AI and data scraping services conspired to illegally scrape copyrighted Reddit content from Google search results

1

. The decision allows Reddit to continue pressing claims that Perplexity and three data scrapers—SerpApi, Oxylabs, and AWMProxy—unlawfully circumvented protective measures to access content for AI training data

3

.

Source: The Verge

Source: The Verge

How the DMCA Lawsuit Differs from Google's Failed Case

The ruling comes less than two weeks after another court dismissed a similar action raised by Google, finding insufficient proof that rights holders like Reddit had authorized the search engine to prevent scraping

1

. Judge Engelmayer distinguished Reddit's case by noting the company went beyond bare allegations, arguing its licensing agreement with Google directly prohibits certain uses of Reddit data now being accessed through anti-circumvention measures employed by third-party scrapers

1

. Reddit successfully argued that when it licenses content to partners like Google, those partners agree to delete posts flagged when users remove content—a promise that millions of monthly post deletions depend upon

1

.

Reddit's Standing to Sue Under Copyright Law

Judge Engelmayer ruled that Reddit has standing to sue for Perplexity's alleged misuse of its users' content, despite the platform not owning the copyrighted material itself

3

. This addresses a key question raised by DMCA experts about whether Reddit, as neither the copyright owner nor the creator of Google's protective technology, could bring such claims

1

. The decision suggests courts may accept that platforms can enforce protections on behalf of their user communities when licensing agreements are in place.

Source: Ars Technica

Source: Ars Technica

What Reddit Seeks in Injunction and Monetary Damages

Reddit has asked the court for multiple forms of relief: an injunction blocking SerpApi and Perplexity AI access to both Reddit and Google websites, another injunction stopping circumvention of Google SearchGuard, and a third preventing the defendants from using previously scraped data

1

. The platform also seeks unspecified monetary damages

3

. Ben Lee, Reddit's chief legal officer, stated the ruling brings the company closer to holding bad actors accountable, emphasizing that Reddit supports responsible access to public content but opposes companies that bypass protections and profit off communities without permission

2

.

Perplexity and SerpApi Vow to Defend Open Internet

Perplexity AI responded defiantly, stating that Reddit's suit claims the right to control access to public web pages it doesn't own, using a security tool it didn't build, on behalf of users it hasn't asked

3

. The company vowed to defend the open internet and predicted victory. SerpApi similarly argued that Google and Reddit are trying to use the DMCA to wall off the open Internet by retroactively claiming control over content they didn't author and don't own

1

. SerpApi's attorney stated the company accesses public search results, not Reddit's platform, and that public information doesn't become protected simply because a platform wants to charge for it

3

.

Implications for AI-Powered Search Engine Licensing

If Reddit wins this AI copyright lawsuit, the platform could be better positioned to force all AI scrapers to enter licensing agreements

1

. Reddit, which features thousands of interest-based subreddit communities, has already licensed its content to Google, OpenAI, and others for AI training

3

. The lawsuit claims Reddit is the most commonly cited source for AI-generated answers to user questions, making control over this data particularly valuable

3

. This case joins many filed by content owners including authors, music labels, and news outlets against tech companies over alleged misuse of copyrighted material for AI systems

3

.

What to Watch as Discovery Proceeds

Judge Engelmayer noted that SerpApi and Perplexity AI may still prove through discovery that Reddit never authorized Google to protect its content in Google search results

1

. SerpApi may also strengthen its defense by proving that publicly accessible content in search results isn't protected by the Copyright Act

1

. The fight appears far from over, with fundamental questions about who controls public web content and how anti-circumvention measures apply to AI training still unresolved. Reddit has also filed a separate data-scraping lawsuit against Anthropic that remains ongoing in California state court

3

.

Today's Top Stories

© 2026 TheOutpost.AI All rights reserved