Google and Reddit's DMCA Scraping Lawsuits Face Legal Reality After Court Defeat
· 3 min read ·
Google has confirmed it will not abandon its legal campaign to prevent AI bots from scraping its search results, even after suffering a major courtroom defeat last week. The search giant's fight has drawn Reddit into a parallel battle, with both companies turning to an unexpected legal framework to push back against the surge in automated data collection.
Google's DMCA Lawsuit Against SerpApi
In December, Google filed a lawsuit against SerpApi, a web scraping service, invoking the Digital Millennium Copyright Act (DMCA). The company accused SerpApi of bypassing its anti-scraping technology and then selling data harvested from Google search results through an unauthorized product marketed as a "Google Search API" software service.
Google argued that its anti-scraping measures existed to safeguard copyrighted material within search results. The company claimed that SerpApi's circumvention jeopardized its relationships with rights holders, including those who license content for the "knowledge panels" that accompany search results for prominent people and entities.
The complaint also alleged that SerpApi's actions violated Google's terms of service and made it impossible for the company to profit from or offset the cost of what it described as "billions" of automated bot searches.
Reddit's Parallel Legal Push
Google's decision to pursue the DMCA theory was reportedly influenced by a similar lawsuit filed by Reddit in October. Reddit sued both SerpApi and Perplexity, a Google competitor, accusing them of scraping Reddit content that appears within Google search results.
Reddit asserted that SerpApi had evaded two separate layers of security: Reddit's own platform-level scraping controls and Google's controls designed to block scraping of Reddit content within search results. When Google announced its own lawsuit, it referenced Reddit's case and described the filing as a "last resort" to combat "malicious scraping" that overrides the choices of rights holders regarding who may access their content.
Legal Experts Question the DMCA Strategy
The use of the DMCA in these cases has raised eyebrows among legal experts, primarily because Google search results themselves cannot be copyrighted. This makes the legal theory behind both lawsuits unusual, since the DMCA was designed to protect copyrighted works from circumvention of access controls.
Meredith Rose, a senior policy counsel with expertise in DMCA matters at the nonprofit public interest group Public Knowledge, told Ars Technica that Google and Reddit appear to be "sort of grasping at whatever tool is available" in response to the rapid escalation of AI scraping over the past three years.
Rose described the companies' application of the DMCA as "bizarre" and "not what the law had sort of contemplated as a use case," while noting that the strategy is "not surprising." She explained that historically, the DMCA has functioned as an effective mechanism to quickly halt disfavored uses of content and compel negotiations around licensing. Given Google's stated objectives, turning to the DMCA may have seemed like an obvious starting point, even if the legal foundation remains uncertain.
As Google presses forward despite its recent court loss, the outcome of these cases could have significant implications for how platforms protect their data in an era where AI companies increasingly rely on large-scale web scraping to train their models. The tension between data accessibility and platform control continues to intensify, and these legal battles may help define the boundaries of what tools companies can legitimately use to restrict access to publicly available online information.
What do you think about Google and Reddit's approach to fighting web scrapers? Should search results and platform data be fair game for AI companies, or do platforms have a right to control who accesses their content? Share this article and join the conversation.