The post Perplexity caught red-handed scraping data, Reddit claims appeared on BitcoinEthereumNews.com. Reddit has sued Perplexity AI for continuing to use Reddit’s content to train its AI model after prior warnings not to scrape the platform’s content.  As AI systems increasingly rely on publicly available online content to train and generate answers, companies like Reddit are trying to draw firm lines over what is considered “public” and “proprietary” data. Reddit’s trap exposes alleged data theft   Reddit has filed a lawsuit against Perplexity, a $20 billion AI company, accusing it of illegally collecting data through its platform. According to court documents filed Wednesday in a Manhattan federal court, Reddit said Perplexity ignored instructions not to scrape its content and continued to use Reddit data to generate AI answers. The complaint says Reddit had explicitly blocked Perplexity from collecting its data, but the AI company’s “answer engine” still produced results containing Reddit content. “The increase was so dramatic that an outside observer hypothesized that the increase was due to Perplexity entering a licensing deal with Reddit,” the lawsuit said. “In truth, there is no license between Perplexity and Reddit.” To prove its suspicion, Reddit designed a clever digital test. It created a “trap” post that could only be found by Google’s search engine. Google has a legitimate content-licensing deal with Reddit, and so any company without such a deal should have been unable to access the post. The company described it as the online equivalent of a “marked bill.” If Perplexity’s system reproduced the contents of that hidden post, Reddit would know it had gone around its safeguards possibly by pulling data through Google’s search results, known as SERPs. Within hours, the supposedly private test post began showing up in responses generated by Perplexity’s AI tool.  “The only way that Perplexity could have obtained that Reddit content and then used it in its ‘answer… The post Perplexity caught red-handed scraping data, Reddit claims appeared on BitcoinEthereumNews.com. Reddit has sued Perplexity AI for continuing to use Reddit’s content to train its AI model after prior warnings not to scrape the platform’s content.  As AI systems increasingly rely on publicly available online content to train and generate answers, companies like Reddit are trying to draw firm lines over what is considered “public” and “proprietary” data. Reddit’s trap exposes alleged data theft   Reddit has filed a lawsuit against Perplexity, a $20 billion AI company, accusing it of illegally collecting data through its platform. According to court documents filed Wednesday in a Manhattan federal court, Reddit said Perplexity ignored instructions not to scrape its content and continued to use Reddit data to generate AI answers. The complaint says Reddit had explicitly blocked Perplexity from collecting its data, but the AI company’s “answer engine” still produced results containing Reddit content. “The increase was so dramatic that an outside observer hypothesized that the increase was due to Perplexity entering a licensing deal with Reddit,” the lawsuit said. “In truth, there is no license between Perplexity and Reddit.” To prove its suspicion, Reddit designed a clever digital test. It created a “trap” post that could only be found by Google’s search engine. Google has a legitimate content-licensing deal with Reddit, and so any company without such a deal should have been unable to access the post. The company described it as the online equivalent of a “marked bill.” If Perplexity’s system reproduced the contents of that hidden post, Reddit would know it had gone around its safeguards possibly by pulling data through Google’s search results, known as SERPs. Within hours, the supposedly private test post began showing up in responses generated by Perplexity’s AI tool.  “The only way that Perplexity could have obtained that Reddit content and then used it in its ‘answer…

Perplexity caught red-handed scraping data, Reddit claims

Reddit has sued Perplexity AI for continuing to use Reddit’s content to train its AI model after prior warnings not to scrape the platform’s content. 

As AI systems increasingly rely on publicly available online content to train and generate answers, companies like Reddit are trying to draw firm lines over what is considered “public” and “proprietary” data.

Reddit’s trap exposes alleged data theft  

Reddit has filed a lawsuit against Perplexity, a $20 billion AI company, accusing it of illegally collecting data through its platform. According to court documents filed Wednesday in a Manhattan federal court, Reddit said Perplexity ignored instructions not to scrape its content and continued to use Reddit data to generate AI answers.

The complaint says Reddit had explicitly blocked Perplexity from collecting its data, but the AI company’s “answer engine” still produced results containing Reddit content. “The increase was so dramatic that an outside observer hypothesized that the increase was due to Perplexity entering a licensing deal with Reddit,” the lawsuit said. “In truth, there is no license between Perplexity and Reddit.”

To prove its suspicion, Reddit designed a clever digital test. It created a “trap” post that could only be found by Google’s search engine. Google has a legitimate content-licensing deal with Reddit, and so any company without such a deal should have been unable to access the post.

The company described it as the online equivalent of a “marked bill.” If Perplexity’s system reproduced the contents of that hidden post, Reddit would know it had gone around its safeguards possibly by pulling data through Google’s search results, known as SERPs.

Within hours, the supposedly private test post began showing up in responses generated by Perplexity’s AI tool. 

“The only way that Perplexity could have obtained that Reddit content and then used it in its ‘answer engine’ is if it and/or its co-defendants scraped Google SERPs,” the lawsuit stated.

Reddit named three data-scraping companies in the suit, Oxylabs UAB, AWM Proxy, and SerpApi. It accused them of helping Perplexity gain unauthorized access to Reddit’s posts, or of selling Reddit’s data to Perplexity.

Reddit’s allegations denied 

Perplexity has rejected Reddit’s allegations. The company’s spokesperson Jesse Dwyer stated that Perplexity “will not tolerate threats against openness and the public interest.” The company also said in a Reddit post after the lawsuit was filed that it “does not train AI models on content.”

Representatives of the other companies named in the lawsuit also issued statements. A spokesperson for SerpApi said it plans to “vigorously defend” itself in court. Oxylabs’ chief governance and strategy officer, Denas Grybauskas, said his company was “shocked and disappointed,” adding that Oxylabs “has always been and will continue to be a pioneer and an industry leader in public data collection.”

In August, Cloudflare, an internet infrastructure company, revealed it had conducted a similar test to see if Perplexity was following web-crawling rules. Cloudflare said it created pages marked with code telling Perplexity’s bots not to access them, but it still found the AI company’s crawlers visiting the restricted pages.

Cloudflare’s CEO, Matthew Prince, made headlines by comparing Perplexity’s behavior to that of “North Korean hackers.” 

“Some supposedly ‘reputable’ AI companies act more like North Korean hackers,” Prince wrote on X. “Time to name, shame, and hard block them.” Reddit’s lawsuit quoted Prince’s remarks as part of its case.

The smartest crypto minds already read our newsletter. Want in? Join them.

Source: https://www.cryptopolitan.com/perplexity-caught-scraping-data-reddit/

Market Opportunity
RedStone Logo
RedStone Price(RED)
$0.2303
$0.2303$0.2303
-2.90%
USD
RedStone (RED) Live Price Chart
Disclaimer: The articles reposted on this site are sourced from public platforms and are provided for informational purposes only. They do not necessarily reflect the views of MEXC. All rights remain with the original authors. If you believe any content infringes on third-party rights, please contact service@support.mexc.com for removal. MEXC makes no guarantees regarding the accuracy, completeness, or timeliness of the content and is not responsible for any actions taken based on the information provided. The content does not constitute financial, legal, or other professional advice, nor should it be considered a recommendation or endorsement by MEXC.

You May Also Like

ZKP Crypto’s $1.7B Presale Changes the Math as ETH Struggles and Dogecoin Searches for Direction!

ZKP Crypto’s $1.7B Presale Changes the Math as ETH Struggles and Dogecoin Searches for Direction!

Uncover why Ethereum prediction remains cautious, Dogecoin price stays sentiment-driven, while ZKP crypto’s $1.7B presale scale positions it as the next crypto
Share
coinlineup2026/01/26 01:00
OpenVPP accused of falsely advertising cooperation with the US government; SEC commissioner clarifies no involvement

OpenVPP accused of falsely advertising cooperation with the US government; SEC commissioner clarifies no involvement

PANews reported on September 17th that on-chain sleuth ZachXBT tweeted that OpenVPP ( $OVPP ) announced this week that it was collaborating with the US government to advance energy tokenization. SEC Commissioner Hester Peirce subsequently responded, stating that the company does not collaborate with or endorse any private crypto projects. The OpenVPP team subsequently hid the response. Several crypto influencers have participated in promoting the project, and the accounts involved have been questioned as typical influencer accounts.
Share
PANews2025/09/17 23:58
How to earn from cloud mining: IeByte’s upgraded auto-cloud mining platform unlocks genuine passive earnings

How to earn from cloud mining: IeByte’s upgraded auto-cloud mining platform unlocks genuine passive earnings

The post How to earn from cloud mining: IeByte’s upgraded auto-cloud mining platform unlocks genuine passive earnings appeared on BitcoinEthereumNews.com. contributor Posted: September 17, 2025 As digital assets continue to reshape global finance, cloud mining has become one of the most effective ways for investors to generate stable passive income. Addressing the growing demand for simplicity, security, and profitability, IeByte has officially upgraded its fully automated cloud mining platform, empowering both beginners and experienced investors to earn Bitcoin, Dogecoin, and other mainstream cryptocurrencies without the need for hardware or technical expertise. Why cloud mining in 2025? Traditional crypto mining requires expensive hardware, high electricity costs, and constant maintenance. In 2025, with blockchain networks becoming more competitive, these barriers have grown even higher. Cloud mining solves this by allowing users to lease professional mining power remotely, eliminating the upfront costs and complexity. IeByte stands at the forefront of this transformation, offering investors a transparent and seamless path to daily earnings. IeByte’s upgraded auto-cloud mining platform With its latest upgrade, IeByte introduces: Full Automation: Mining contracts can be activated in just one click, with all processes handled by IeByte’s servers. Enhanced Security: Bank-grade encryption, cold wallets, and real-time monitoring protect every transaction. Scalable Options: From starter packages to high-level investment contracts, investors can choose the plan that matches their goals. Global Reach: Already trusted by users in over 100 countries. Mining contracts for 2025 IeByte offers a wide range of contracts tailored for every investor level. From entry-level plans with daily returns to premium high-yield packages, the platform ensures maximum accessibility. Contract Type Duration Price Daily Reward Total Earnings (Principal + Profit) Starter Contract 1 Day $200 $6 $200 + $6 + $10 bonus Bronze Basic Contract 2 Days $500 $13.5 $500 + $27 Bronze Basic Contract 3 Days $1,200 $36 $1,200 + $108 Silver Advanced Contract 1 Day $5,000 $175 $5,000 + $175 Silver Advanced Contract 2 Days $8,000 $320 $8,000 + $640 Silver…
Share
BitcoinEthereumNews2025/09/17 23:48