Reddit sues Anthropic, alleging it accessed Reddit 100K+ times after saying it had stopped; Reddit has reached user data licensing deals with OpenAI and Google
For a while it looked like Reddit was going to build a big business selling access to user data to LLM companies but besides Google they didn't get many takers it seems. — So here come the lawyers. Doll / @dollspace.gay : A couple of clueless quotes on this one but heres the thing. — You all gave reddit licence to do whatever they wanted with your posts when you signed up for the site. A deal most people took without thinking. [embedded post] @vortexegg.com : This is kind of funny given how many reddit posts appear to now be AI generated. Licensing slop back to the slop mongers to train more models on [embedded post] X: Andrew Curran / @andrewcurran_ : Google pays reddit $60 million a year to train off their data, OpenAI allegedly pays $70 million a year. Jamin Ball / @jaminball : You don't have an AI strategy without a Data strategy: Reddit / Anthropic: https://www.cnbc.com/... Slack updates API terms of service: https://slack.com/... Anthropic cuts off Windsurf: https://x.com/... Shirin Ghaffary / @shiringhaffary : Reddit is suing Anthropic over use of its data. Reddit's chief legal officer told us that until recently, they had been in talks about reaching a licensing agreement; lawsuit is a final option to force Anthropic to bargaining table. https://www.bloomberg.com/... @xlr8harder : Accessing the site 100,00 times isn't actually all that much. This could easily be a single misconfigured web crawler. Annoying, but it hardly seems worthy of a law suit. @andrewarruda : “For CoCounsel to be trustworthy and immediately useful for practicing attorneys, it needs to cite its work. We first built this ourselves, but it was really hard to build and maintain.” 👀 Fair use for me but not for thee. [image] Kyle ‘esSOBi’ Stone / @essobi : ....... It's starting? Good. Let them eat each other while we go make a Creative Commons model. @luke_metro : The YC mob has put out a hit Alex J. Champandard / @alexjc : This will probably get settled; Reddit just suing to increase the pressure? The AI companies, including Anthropic, don't want to set a new precedent that they must respect Terms of Service... Prior claims (e.g. Genius v Google) were deemed to be preempted by © law and dismissed. LinkedIn: Emil Protalinski / LinkedIn : Emil Protalinski's Post Forums: r/technology : Reddit sues Anthropic for allegedly not paying for training data r/ArtificialInteligence : Reddit sues Anthropic for allegedly not paying for training data r/SEO : Reddit sues AI startup Anthropic for breach of contract, ‘unfair competition’ over training r/technology : Reddit sues Anthropic, alleging its bots accessed Reddit more than 100,000 times since last July r/artificial : Reddit sues Anthropic, alleging its bots accessed Reddit more than 100,000 times since last July r/singularity : Reddit Sues Anthropic, Alleges Unauthorized Use of Site's Data r/wallstreetbets : Reddit Sues Anthropic, Alleges Unauthorized Use of Site's Data See also Mediagazer
Context & Ripple Effects
Reddit had already begun turning its corpus into a licensed AI input: Google received access through Reddit's Data API, and Reddit later partnered with OpenAI to bring its content into OpenAI tools. The Anthropic suit tests whether that commercial model can be enforced against model developers that do not reach an agreement.
The dispute also follows regulatory attention to Reddit's data-licensing approach, including an FTC inquiry into those arrangements. That makes the case consequential beyond a single alleged scraper: it concerns the boundary between platform-controlled access and AI training demand.
First-order effects
- Reddit and Anthropic move from licensing talks into litigation, with Reddit alleging Anthropic continued accessing the site more than 100,000 times after saying it had stopped.
- The suit puts Anthropic's alleged collection practices under legal challenge while reinforcing Reddit's position that AI access should run through a commercial agreement.
Second-order effects
- Reddit's existing API partners, including the OpenAI content partnership, gain a clearer contrast between licensed access and access Reddit characterizes as unauthorized.
- Other AI developers seeking Reddit material may face greater pressure to negotiate API terms rather than rely on web collection, strengthening Reddit's leverage in those talks.
Third-order effects
- If platforms can consistently convert alleged scraping into enforceable claims, training-data access could shift further from open-web collection toward paid, controlled data channels.
- The unresolved question is whether courts and regulators will validate those boundaries; the answer will shape how much negotiating power content platforms hold over model-training inputs.
The trend: AI developers' demand for high-quality conversational data is turning platform data access into a negotiated, litigated market rather than a presumed byproduct of the open web.