Reddit v. Anthropic: The Lawsuit That Exposes a Big AI Outlier
The AI darling that championed ethical AI also refused to pay for content. Now Reddit is calling it out.
Of the “Frontier 5” AI labs, only one hasn’t paid publishers for their content.
That outlier? Anthropic.
I’ve been tracking content licensing deals for months (see “5 Takeaways from the AI Content Licensing Deals I'm Tracking”) — and the other four frontier labs have all written checks to publishers:
OpenAI: Pays News Corp ($50M/yr), Shutterstock ($37.5M), Dotdash Meredith ($16M), and Reddit (undisclosed)
Google: Pays Reddit ($60M), Shutterstock ($37.5M), and AP (amount unknown)
Meta: Pays Shutterstock ($37.5M) and Reuters (amount unknown)
xAI: Pays X.com (Twitter) $ hundreds of millions (even though Elon owns both)
Even Amazon (not usually in the licensing mix) struck a deal with The New York Times. That amount wasn’t disclosed but I’d be shocked if it wasn’t in the tens of millions per year.
But Anthropic? Zip. Zero.
That’s why Reddit’s new lawsuit against Anthropic caught my eye.
It cracked open a chance to dive deeper—and what I found shows just how far Anthropic sits outside the AI-content licensing herd.
The Reddit Lawsuit
Reddit’s lawsuit against Anthropic pulls no punches—and it’s not your usual copyright case.
Here are the highlights:
Over 100,000 bot hits – Reddit claims Anthropic’s bots accessed the site more than 100,000 times, even after saying in July 2024 that they had blocked Reddit scraping. Audit logs say otherwise.
Claude admits it – Reddit shows a screenshot of Claude (Anthropic’s chatbot) admitting it was trained on Reddit data.
Anthropic’s CEO confirms it – In a 2021 research paper, CEO Dario Amodei explicitly says their models were trained on Reddit, Wikipedia, and Stack Exchange.
This isn’t a copyright fight – Reddit’s not suing over IP. They’re going after Anthropic for breach of contract, platform abuse, and unfair competition.
This isn’t just legal posturing. Reddit is accusing Anthropic of quietly scraping their content while OpenAI and Google paid for it.
The message is clear: if the others cut a check, why didn’t you?
Anthropic Admits It’s Trained on Reddit Data
A December 2021 paper co-authored by CEO Dario Amodei plainly states Anthropic trained on Reddit, Stack Exchange, and Wikipedia. The Reddit lawsuit cites it as proof.
One wrinkle: the paper references Pushshift.io—a third-party Reddit dataset—so the legal question is whether Anthropic scraped Reddit directly or relied on someone else’s scrape.
Either way, Reddit’s point is clear: You used our data. You didn’t pay.
The Reddit Case is Not About Copyright
Unlike other high-profile lawsuits—like The New York Times v. OpenAI or Getty v. Stability AI—Reddit isn’t suing over copyright.
Instead, Reddit is coming after Anthropic with five different claims that focus on contracts, platform abuse, and unfair competition:
Breach of Contract – Anthropic broke Reddit’s User Agreement by scraping the site without permission—and then monetized the data.
Unjust Enrichment – Reddit says Anthropic made billions using Reddit content without paying a dime.
Trespass to Chattels – Anthropic’s bots overloaded Reddit’s servers—without consent.
Tortious Interference – Anthropic allegedly violated Reddit’s agreements with its users by harvesting their content without following the rules.
Unfair Competition – By skipping the licensing fees others paid, Anthropic gained an edge that Reddit says is illegal.
One thing Reddit makes clear: OpenAI and Google played by the rules and paid for Reddit content.
Anthropic didn’t—and Reddit wants the courts to call that out.
Reddit: There’s an “Established Market for Licensing Content”
Reddit’s argument is simple: we sell our data—and the market has spoken.
In the lawsuit, Reddit points out that it has an established business around licensing its content to tech companies.
And not just small deals—major players are already paying:
“OpenAI, Google, Sprinklr, and Cision have all agreed to formal licensing agreements with Reddit in exchange for lawful access to Reddit public content.” — Section 55, Reddit v. Anthropic
These aren’t token partnerships. Reddit calls them proof of the “immense value” of its user-generated content.
By refusing to pay—and scraping anyway—Anthropic positioned itself as the odd one out.
That’s a key part of Reddit’s argument: if its biggest competitors can respect the rules and license data, why can’t Anthropic?
Reddit’s “Uphill Battle” in the Case
Reddit may have a strong story—but winning in court won’t be easy.
An intellectual property attorney I spoke with flagged three big hurdles Reddit faces:
Copyright Law Might Block Their Claims
Reddit didn’t sue for copyright—probably because it doesn’t own the rights to most user posts. But the lawsuit is still about Anthropic copying content. If that’s the case, a judge could say only federal copyright law applies—and toss out Reddit’s state-level claims.The Contract Might Not Count
Reddit says Anthropic broke its User Agreement by scraping the site. But Anthropic never signed anything or clicked “I agree.” The rules were just posted online—what courts call a “browsewrap” agreement. Anthropic could argue: “We never agreed to your terms, so we didn’t break them.”No Proof of Server Harm
Reddit also claims that Anthropic burdened its servers. But under California law, that only holds up if it caused real damage—like slowing or crashing systems. Just hitting the site a lot? That’s not enough.
Anthropic's Silence
While other AI giants are cutting content deals or at least talking about them, Anthropic has stayed mostly silent.
That’s a sharp contrast to how loudly they promote their signature approach: “Constitutional AI.”
According to Anthropic, this method bakes values like safety, honesty, and harmlessness into its models.
“The constitution guides the model to take on the normative behavior described in the constitution—avoiding toxic or discriminatory outputs, avoiding helping a human engage in illegal or unethical activities...”
— Anthropic.com’s Constitution
That kind of language may come back to haunt them.
Reddit’s lawyers are likely circling phrases like “ethical” and “harmless.”
But when it comes to how Anthropic sources its training data—especially from content creators—they’ve said next to nothing.
No strategy. No philosophy. Just silence.
What Anthropic Says About Paying for Content
Short answer: Nothing.
Unlike OpenAI’s Sam Altman who acknowledges the value of publisher relationships—Anthropic’s top execs have said zero about paying content creators.
No vision. No blog post. No panel appearance explaining how they plan to build a sustainable data ecosystem.
Not from CEO Dario Amodei. And not from any of the other 6 co-founders.
If Anthropic has a philosophy around licensing content, they’re keeping it to themselves.
Anthropic Argues “Fair Use”
While Anthropic hasn’t said much about licensing content, they’ve been crystal clear about one thing:
They believe training on copyrighted data—without paying—is protected by “fair use.”
Here’s what CEO Dario Amodei told The New York Times’ Ezra Klein:
“I think everyone agrees the models shouldn’t be verbatim outputting copyrighted content... our position [is] that this is sufficiently transformative... that this is fair use.”
— Ezra Klein Interview, April 2024
Anthropic doubled down on that stance in its response to a lawsuit from book authors:
Making “intermediate” copies to study how words and concepts relate is “a quintessential fair use.”
Translation: As long as the model doesn’t spit out large, word-for-word passages, training on the content is fair game.
That legal theory may hold up in court—but it puts Anthropic even further out of step with the rest of the industry.
How Anthropic Says it Gets its Data
Anthropic doesn’t reveal much—but they’ve shared a few lines about where Claude gets its training fuel.
Here’s what their website says (last updated May 2025):
“Claude Opus 4 and Claude Sonnet 4 were trained on a proprietary mix of publicly available information on the Internet as of March 2025, as well as non‑public data from third parties, data provided by data‑labeling services and paid contractors, data from Claude users who have opted in to have their data used for training, and data we generated internally at Anthropic.”
— Anthropic System Card, March 2025
Translation: They won’t name names.
The Closest Anthropic Has Come to a Deal with a Content Publisher
In January 2025, Anthropic made headlines for settling with a group of major music publishers—including Universal Music.
But let’s be clear: this was not a licensing deal.
Anthropic didn’t pay Universal.
Instead, they agreed to set up guardrails—measures to help prevent Claude from spitting out copyrighted lyrics owned by the music publishers.
“It’s more a ceasefire than a peace treaty.” — Tech Policy Press
Still, it’s progress. It shows that Anthropic can come to the table—even if they’re not cutting checks yet.
The big question now: will the Reddit lawsuit force a similar compromise?
The Bigger Story? Anthropic’s Brand Trust is at Risk
Anthropic didn’t just market itself as another AI lab—it positioned itself as the ethical one.
They coined phrases like “Constitutional AI”, and promised to build models that are safe, honest, and harmless.
They wrote long essays about aligning AI with human values. They emphasized safety over speed. Alignment over dominance. Ethics over scale.
But Reddit’s lawsuit puts that entire narrative at risk.
Can a company claim the moral high ground while quietly training on content it didn’t pay for?
That’s not just a legal question—it’s a brand trust issue.
And the cracks may already be showing.
Anthropic Kills Claude Explains Project
Last week, Anthropic launched Claude Explains, an AI-written blog meant to help the public understand Claude’s thinking.
The idea sounded noble: transparency, interpretability, AI explaining AI.
But within days, the company shut it down.
Why? TechCrunch reports that people caught Claude citing fake sources, mixing facts with fiction, and—ironically—breaking the very trust the project was supposed to build. — Source: “Anthropic’s AI-generated blog dies an early death.”
So now you have an AI company under fire for quietly scraping data… and for backtracking on its biggest transparency initiative.
The legal risks are real. But the brand risk might cut even deeper.
AI’s “Original Sin”

The Atlantic CEO Nick Thompson sums up the case perfectly:
“AI companies, of course, went and scraped the whole Internet and did not compensate the sites they scraped... I believe this is the original sin of the AI industry.”
The Atlantic licensed its content to OpenAI in May 2024 for an undisclosed amount.
Thompson thinks this will get worked out—but it’s going to be messy. He sees three ways forward:
Court cases (like Reddit v. Anthropic)
Deals (like OpenAI–News Corp, Google–Reddit)
New laws (some are already on the way)
What’s needed, according to Thompson, is a “fair exchange of value”—just like the old Google/publisher dynamic: Google got snippets; publishers got traffic.
I agree.
He adds that Anthropic, with its heavy focus on ethics and safety, should be a leader in figuring this out.
“They’ve tried to make the world a better, more sustainable place... we’ll see what happens now that Reddit has started this big fight.”
Soon, Anthropic Might Have to Reveal Its Datasets
While the courts sort things out, lawmakers are starting to move.
At least three U.S. states—Tennessee, Colorado, and California—have passed laws that push for AI data transparency.
These laws could require companies like Anthropic to disclose what data they trained on—at least within those states.
It’s still unclear how enforceable these laws are—or whether any AI lab has actually complied yet.
But the signal is clear: the era of “secret datasets” may be ending.
And for a company like Anthropic, which is already under fire for silence and scraping, that could get very uncomfortable, very fast.
Final Takeaways
For Content Executives:
The AI licensing market is real—Reddit just added proof to it ( ( in court.
If your content trains models, it’s an asset—track where it’s used.
Don’t wait for AI companies to come to you; start the licensing convo.
For AI Executives:
You're facing a clear fork: license content or roll the dice on Fair Use (Anthropic is choosing the latter).
Whichever path you take, remember that Trust transcends all strategy.
Licensing is becoming the norm. Paying now may be cheaper than litigating later.
Thanks for reading!
Rob Kelly, Creator & Host of Media & the Machine
p.s.: How to reach me:
If you want to reach me, the best way is to subscribe below and reply to my emails. There’s a free option and I read every email.




