Reddit is currently engaged in a complex legal battle with web scrapers and a tense commercial standoff with Google. While fighting a DMCA lawsuit against SerpApi and Perplexity AI, Reddit is simultaneously threatening to block Google’s AI models from its content as the two companies renegotiate a data-licensing deal.
The tension between social platforms and AI developers has reached a boiling point. Reddit, whose vast library of human conversation fuels much of the web’s search results, is fighting on two fronts: in the courtroom over copyright protections and in the boardroom over the price of its data.
The DMCA Dispute with SerpApi and Perplexity AI
Reddit’s legal strategy has taken an unconventional turn in its fight against SerpApi and Perplexity AI. The platform is attempting to use the Digital Millennium Copyright Act (DMCA) to stop these entities from scraping its content via Google search results.
The case is fraught with technical hurdles. Meredith Rose, a DMCA expert with Public Knowledge, suggested that both Google and Reddit appear to be somewhat grasping at whatever tool is available to combat the surge of AI scraping seen over the last three years. Rose pointed out a significant gap in Reddit’s standing, noting that to bring a lawsuit under the DMCA, a party must be the copyright owner, an exclusive licensee, or the entity deploying the technological protection measure.
Meredith Rose of Public Knowledge noted that the judge in the Google case stated that standing to bring a DMCA lawsuit requires being the copyright owner, the exclusive licensee, or the person deploying and manufacturing the technological protection measure at issue, and that Reddit does not meet any of those criteria.
Despite these challenges, the fight continues. While the court tossed Reddit’s claims of unfair competition and unjust enrichment because they were preempted by the Copyright Act, the DMCA claim remains alive. Rose noted that DMCA rulings often depend on vibes
, making the final outcome difficult to predict.
SerpApi’s Defense of Public Search Results
SerpApi is not conceding. The company argues that it is accessing information that is already public, rather than infiltrating Reddit’s private platform. Jeff Homrig, a lawyer for SerpApi, maintains that the facts support their position.
Jeff Homrig, Lawyer for SerpApi, stated that they remain confident in their position, noting that the court has decided to hear the facts which are on their side. He added that SerpApi accesses public search results rather than Reddit’s platform, and that public information does not become protected simply because a platform wants to charge for it, and he looks forward to making that case.
The legal pivot now rests on discovery. SerpApi and Perplexity AI may attempt to prove that Reddit never actually authorized Google to protect its content within search results. Furthermore, a footnote in a legal opinion suggests that if the defendants can prove publicly accessible content in Google search results is not protected by the Copyright Act, their defense will strengthen significantly.
The Google Licensing Conflict
While the lawsuit plays out, Reddit is leveraging its relationship with Google to protect its traffic. In 2024, the two companies signed a deal, which allowed Google to use Reddit’s message boards to train its AI models.
That agreement is now under scrutiny. Reports indicate that Reddit officials are considering blocking Google’s access as they negotiate the renewal of their data-licensing terms. This shift is driven by traffic cannibalization. Google’s AI-generated answer boxes now provide direct responses to user queries on the search page, removing the need for users to click through to external forums like Reddit.
This creates a critical financial paradox for Reddit: the very data that makes Google’s search results valuable is being used by Google’s AI to ensure users never actually visit Reddit. Because Reddit relies on visitor traffic to sell advertising, the AI shift is directly cutting into its revenue potential.
A Broader War Over Generative AI Data
Reddit is not alone in this struggle. Other media giants, including Reuters, Politico, The Economist, USA Today, and People Inc., are also reevaluating their partnerships with Google under similar pressures. The industry is shifting from viewing data licensing as a passive revenue stream to treating it as a primary battlefield for the future of web traffic.
The urgency is compounded by the scale of automation. Data from Cloudflare shows that automated bot activity now accounts for more than 50% of global web traffic. This has made platforms increasingly protective of “real” human conversation, which has become the essential lifeline for generative AI.
The stakes for Reddit are immediate. The market has already reacted to these tensions; the company’s shares saw a 5.8% drop in pre-market trading following reports of the threats to block Google’s AI training access.
