Judge Allows Reddit's DMCA Claims Against AI Data Scrapers

3 min readSources: LegalTech News

A federal judge permits Reddit to proceed with DMCA claims against AI scraping firms.

Why it matters: This ruling could shape how courts enforce copyright protections on content used for AI training. Legal teams must watch evolving AI liability and content rights issues carefully.

  • Reddit sued Perplexity AI, SerpApi, Oxylabs UAB, and AWMProxy in October 2025 for unauthorized scraping of Reddit content.
  • SerpApi argued in April 2026 that Reddit lacked standing due to a non-exclusive license over user content.
  • On July 31, 2026, a Manhattan federal judge denied motions to dismiss, allowing Reddit's DMCA claims to proceed.
  • The case highlights growing legal tensions over AI training data and DMCA protections.

On October 22, 2025, Reddit filed suit in the U.S. District Court for the Southern District of New York against AI companies Perplexity AI, SerpApi, Oxylabs UAB, and AWMProxy. The complaint alleges these defendants scraped Reddit's content without authorization by bypassing technological safeguards and reselling the data, thus violating the Digital Millennium Copyright Act (DMCA).

In April 2026, SerpApi moved to dismiss the complaint, contending that Reddit lacked standing because it holds only a non-exclusive license to user-generated content rather than full ownership. This argument raised questions about who holds copyright interests in platform content.

On July 31, 2026, a federal judge in Manhattan rejected the defendants’ motions to dismiss, allowing Reddit’s DMCA claims to move forward. Although the judge's detailed reasoning was not publicly released, the decision is significant as it allows legal scrutiny of AI scraping practices under copyright law.

Ben Lee, Reddit’s Chief Legal Officer, described the case as combatting what he termed an industrial-scale “data laundering” economy, where user-generated content is exploited for AI training without proper authorization. In contrast, SerpApi’s founder, Julien Khaleghy, asserted that Reddit is attempting to monetize user content without consent, emphasizing that content creators retain ownership over their posts rather than Reddit.

This lawsuit highlights intensifying legal challenges as AI companies harvest massive online data sets. The case is a bellwether for how courts may apply the DMCA to unauthorized content extraction for AI model training, a central issue in ongoing debates regarding AI liability, copyright enforcement, and platform control over content rights.

By the numbers:

  • October 22, 2025 — Date Reddit sued AI scraping firms in SDNY.
  • April 2026 — SerpApi filed motion to dismiss Reddit’s complaint.
  • July 31, 2026 — Judge denied motions, allowing Reddit’s DMCA claims to proceed.

Yes, but: The judge’s ruling does not resolve all questions about ownership and standing; further litigation will clarify these issues.

What's next: Discovery and potential settlement talks are expected as this case proceeds in SDNY, with broader implications for AI data usage.