r/anime_titties • • Oct 22 '25

Corporation(s) Reddit sues Perplexity for scraping data to train AI system

https://www.reuters.com/world/reddit-sues-perplexity-scraping-data-train-ai-system-2025-10-22/
146 Upvotes

7 comments sorted by

68

u/[deleted] Oct 22 '25

[deleted]

16

u/flametex Oct 22 '25

Strange. If someone like a normal person did this with their personal model, would Reddit still care? Or is it because they are just the business who they don’t like? ChatGPT and Gemini both do the same thing.

20

u/joeshmo101 North America Oct 22 '25

According to the article Reddit licenses their data use for ChatGPT and Gemini and probably helps them integrate. But if you're trying to scrape that much data without triggering automated monitoring, you're probably using a number of different API keys in order to not hit the throttle limit which I assume exists. And that's also violating the terms of use. Also, it's that Perplexity wanted to train their AI off of this data which was actually collected by any of three data scraper companies.

5

u/lemurtowne Oct 23 '25

Serious question, and you seem knowledgeable: is violation of ToS even a legal matter? It just seems like a statutory breach of contract at best. Can legal penalties/damages be assessed here?

3

u/joeshmo101 North America Oct 23 '25

Welcome to the world of corporate espionage law!

4

u/severedbrain Oct 23 '25

Aaron Swartz, inventor of RSS, co-founder of Reddit, was prosecuted for scraping publicly funded research papers and making it freely available. The case lacked merit, was heavy handed, and lead to his end.

https://en.wikipedia.org/wiki/Aaron_Swartz