Reddit's Data-Scraping Lawsuit Against Perplexity AI Moves Forward
A Manhattan federal judge rejected most of Perplexity AI's bid to dismiss Reddit's lawsuit over alleged data scraping for AI-powered search.
Why it matters: AI search depends on web-scale content access, while platforms increasingly treat their archives as licensed datasets.
Reddit may continue pursuing its lawsuit accusing Perplexity AI of obtaining platform content without permission for use in an AI-powered search engine.
A federal judge in Manhattan rejected most of Perplexity's request to dismiss the case. The ruling allows Reddit to proceed with claims that Perplexity and several data-scraping companies circumvented technical protections to collect Reddit content.
The decision does not determine that Perplexity violated the law. It means Reddit's core allegations can continue into later stages of litigation.
Reddit alleges that scraping companies collected its content from search results and supplied material to Perplexity, which does not have a Reddit licensing agreement. Perplexity disputes the claims and argues that Reddit is trying to control access to publicly available pages and user-created content.
The case matters because Reddit has licensed content to companies including Google and OpenAI. Those agreements make control over its archive increasingly important as AI developers compete for high-quality conversational material.
The lawsuit illustrates a broader shift in the economics of the open web. Platforms are no longer treating automated access only as a technical issue. They increasingly view their archives as commercial datasets that AI companies must license.