Home AI Tools Blogs AI News About Us Contact Us
➕ Submit AI Tools ✍️ Write for Us
Home AI News Reddit’s AI Copyright Suit Against Perplexity Gets Green Light
Share

Reddit’s AI Copyright Suit Against Perplexity Gets Green Light

Arbaz Khan
AI News Editor & Researcher
Aug 15, 2026
3 min read
AI News

A Manhattan federal judge refused to dismiss Reddit’s data scraping lawsuit against Perplexity AI in early August 2026. U.S. District Judge Paul Engelmayer ruled that Reddit can pursue core claims of unlawful circumvention and conspiracy. This decision keeps the legal battle over unauthorized AI training data completely alive.

The Core Anti Circumvention Claims

Reddit alleges that Perplexity and three specific data scraping companies deliberately bypassed technical barriers. They extracted content from billions of search results without permission or compensation. The defendants include Texas based SerpApi, Lithuania based Oxylabs, and Russia based AWMProxy.

Honestly, most people get this wrong. They assume this is just a standard intellectual property dispute. The actual fight centers on whether circumventing a platform’s digital access controls creates legal liability on its own.

Judge Engelmayer dismissed several secondary claims but allowed the central allegations to move forward. Reddit successfully argued it has standing to sue over the misuse of posts created by its actual users. Digital gates now carry real legal weight.

The Threat to AI Business Models

If bypassing access controls creates immediate liability, the entire AI data pipeline faces massive legal exposure. Perplexity strongly denies the allegations. The company argues Reddit is trying to control access to publicly available web pages on behalf of users who never authorized a lawsuit.

I have observed that frontier AI labs rely heavily on third party scrapers to feed their models. We track these legal battles closely, just like the Suno GEMA German copyright ruling. Content owners now possess another legal lever to protect their proprietary data.

These are the exact practices Reddit accuses the defendants of executing.

  • Masking their identities and locations to scrape content indirectly.
  • Ignoring formal cease and desist notices sent to the company.
  • Purchasing stolen data from middlemen instead of signing lawful agreements.
  • Bypassing technological protections built by both Google and Reddit.

Why Licensing Agreements Matter Now

Reddit built a highly lucrative licensing business by striking deals with Google and OpenAI. When competitors scrape the exact same data around those access controls, it destroys the financial value of those agreements.

In my practical testing of language models, high quality human conversations train agents significantly faster than synthetic data. AI companies desperately need this raw material. If you run a company navigating these shifts, review the top AI tools for business to make sure your data vendors operate legally.

Perplexity claims it will continue defending the open internet. But this court ruling signals that tech startups can no longer treat technical barriers as mere suggestions. The Southern District of New York just declared that breaking digital locks carries massive consequences.

FAQs Reddit and Perplexity Lawsuit

Why is Reddit suing Perplexity AI?

Reddit is suing Perplexity AI for allegedly bypassing technical protections to scrape user data without permission. Reddit claims Perplexity worked with data scraping companies to steal content for its search engine.

Did a judge rule against Perplexity AI?

A Manhattan federal judge refused to dismiss the core of Reddit’s lawsuit. While the judge has not determined if Perplexity broke the law, he allowed the central claims of unlawful circumvention to proceed.

Who else is named in the Reddit scraping lawsuit?

The lawsuit names three data scraping companies alongside Perplexity. These include SerpApi in Texas, Oxylabs in Lithuania, and AWMProxy in Russia.

How does this affect AI training data?

The court ruling suggests that circumventing a website’s access controls to gather training data can create legal liability. This strengthens the position of publishers who want AI companies to pay for data licensing agreements.

Arbaz Khan

Arbaz Khan is a Full-Stack SEO Expert and AI Tools Reviewer at GuideAITools. With 2+ years of hands-on experience in Technical SEO, On-Page, Off-Page, Semantic SEO, AEO, and GEO, he helps businesses rank higher and stay ahead in the AI era. At GuideAITools, Arbaz tests, reviews, and compares AI tools across multiple categories from Audio and Video to Business, Marketing, and Productivity to deliver objective, research-backed content for professionals and beginners alike.

Scroll to Top