Skip to content
October 26, 2025Cryptopolitan logoCryptopolitan

Perplexity caught red-handed scraping data, Reddit claims

Reddit has sued Perplexity AI for continuing to use Reddit’s content to train its AI model after prior warnings not to scrape the platform’s ￰1￱ AI systems increasingly rely on publicly available online content to train and generate answers, companies like Reddit are trying to draw firm lines over what is considered “public” and “proprietary” data. Reddit’s trap exposes alleged data theft Reddit has filed a lawsuit against Perplexity, a $20 billion AI company, accusing it of illegally collecting data through its ￰2￱ to court documents filed Wednesday in a Manhattan federal court, Reddit said Perplexity ignored instructions not to scrape its content and continued to use Reddit data to generate AI ￰3￱ complaint says Reddit had explicitly blocked Perplexity from collecting its data, but the AI company’s “answer engine” still produced results containing Reddit content.

“The increase was so dramatic that an outside observer hypothesized that the increase was due to Perplexity entering a licensing deal with Reddit,” the lawsuit said. “In truth, there is no license between Perplexity and Reddit.” To prove its suspicion, Reddit designed a clever digital ￰4￱ created a “trap” post that could only be found by Google’s search ￰5￱ has a legitimate content-licensing deal with Reddit, and so any company without such a deal should have been unable to access the ￰6￱ company described it as the online equivalent of a “marked bill.” If Perplexity’s system reproduced the contents of that hidden post, Reddit would know it had gone around its safeguards possibly by pulling data through Google’s search results, known as ￰7￱ hours, the supposedly private test post began showing up in responses generated by Perplexity’s AI tool.

“The only way that Perplexity could have obtained that Reddit content and then used it in its ‘answer engine’ is if it and/or its co-defendants scraped Google SERPs,” the lawsuit ￰8￱ named three data-scraping companies in the suit, Oxylabs UAB, AWM Proxy, and ￰9￱ accused them of helping Perplexity gain unauthorized access to Reddit’s posts, or of selling Reddit’s data to Perplexity. Reddit’s allegations denied Perplexity has rejected Reddit’s ￰10￱ company’s spokesperson Jesse Dwyer stated that Perplexity “will not tolerate threats against openness and the public interest.” The company also said in a Reddit post after the lawsuit was filed that it “does not train AI models on content.” Representatives of the other companies named in the lawsuit also issued statements.

A spokesperson for SerpApi said it plans to “vigorously defend” itself in court. Oxylabs’ chief governance and strategy officer, Denas Grybauskas, said his company was “shocked and disappointed,” adding that Oxylabs “has always been and will continue to be a pioneer and an industry leader in public data collection.” In August, Cloudflare, an internet infrastructure company, revealed it had conducted a similar test to see if Perplexity was following web-crawling ￰11￱ said it created pages marked with code telling Perplexity’s bots not to access them, but it still found the AI company’s crawlers visiting the restricted pages. Cloudflare’s CEO, Matthew Prince, made headlines by comparing Perplexity’s behavior to that of “North Korean hackers.” Some supposedly “reputable” AI companies act more like North Korean ￰12￱ to name, shame, and hard block them. ￰0￱ — Matthew Prince 🌥 (@eastdakota) August 4, 2025 “Some supposedly ‘reputable’ AI companies act more like North Korean hackers,” Prince wrote on X.

“Time to name, shame, and hard block them.” Reddit’s lawsuit quoted Prince’s remarks as part of its ￰13￱ smartest crypto minds already read our ￰14￱ in? Join them .

Cryptopolitan logo
Cryptopolitan

Latest news and analysis from Cryptopolitan

Revolutionary Tokenized Stocks: StableStock Unleashes $10M in Digital Assets

Revolutionary Tokenized Stocks: StableStock Unleashes $10M in Digital Assets

BitcoinWorld Revolutionary Tokenized Stocks: StableStock Unleashes $10M in Digital Assets The world of digital finance is buzzing with innovation, and a new player, StableStock, is making waves with i...

Bitcoin World logoBitcoin World
1 min
XRP ETF Hits $100 Million Milestone, Hinting at Broader Institutional Adoption

XRP ETF Hits $100 Million Milestone, Hinting at Broader Institutional Adoption

The REX-Osprey XRP ETF has surpassed $100 million in assets under management within a month of its U.S. debut, signaling strong institutional interest in XRP and its role in mainstream...

CoinOtag logoCoinOtag
1 min
XRP Army Praised On CNBC: “There’s Enormous Interest In XRP”

XRP Army Praised On CNBC: “There’s Enormous Interest In XRP”

During an interview on CNBC, Teucrium President and CEO Sal Gilbertie discussed the company’s leveraged XRP fund and the level of enthusiasm surrounding it. Gilbertie described investor participation ...

TimesTabloid logoTimesTabloid
1 min