Reddit - Policy Change
Executive Summary
Reddit filed a lawsuit against Perplexity AI and three data scraping companies, alleging they illegally harvested millions of user comments from Reddit by collecting the data through Google search results and bypassing Reddit's protections. The lawsuit claims Perplexity purchased scraped Reddit data without authorization or compensation, violating copyright laws and unfair competition rules. This is Reddit's second such case this year, reflecting growing tensions over who controls valuable us...
What Happened
On October 22, 2025, Reddit filed a lawsuit in the U.S. District Court for the Southern District of New York against Perplexity AI and three data scraping companies: Oxylabs UAB, SerpApi, and AWMProxy. The complaint alleges these firms bypassed Reddit's technological protections by harvesting millions of user comments through Google search results, then sold this data to AI companies without Reddit's authorization or compensation. Reddit claims Perplexity purchased scraped Reddit data from at least one of these scraping firms, violating copyright laws and unfair competition rules.
Who Is Affected
Reddit's more than 100 million daily users and 416 million weekly visitors are affected, as their posts and comment threads were allegedly collected and resold without their knowledge or consent. Users who participated in Reddit's public forums had their contributions - ranging from advice and debate to personal anecdotes - harvested as training material for commercial AI systems. The unauthorized data collection impacts anyone who has posted on Reddit, as their human-generated content was reportedly exploited for profit by third parties.
Why It Matters
This case represents a critical test of who controls valuable user-generated content in the AI era and could establish legal precedents for data ownership and scraping practices. Reddit has invested tens of millions of dollars in anti-scraping technology and negotiated paid licensing deals with companies like Google and OpenAI, making unauthorized harvesting a direct threat to its business model. The lawsuit is Reddit's second such action in 2025, following a similar case against Anthropic in June, signaling an industry-wide conflict over whether AI companies can freely use public online conversations or must obtain permission and pay for access.
What You Should Do
Reddit users should understand that their public posts and comments may be collected by third parties despite platform protections, so avoid sharing sensitive personal information in forum discussions. Review your Reddit privacy settings and consider limiting the visibility of past posts if you're concerned about data harvesting. If you want to support platforms that enforce data rights, stay informed about how Reddit and similar sites handle licensing agreements with AI companies. Users cannot directly prevent past scraping but can be more cautious about what they share publicly going forward.
Summary generated from verified sources and reviewed before publication. How we summarize.