Anthropic has reached a $1.5 billion copyright settlement with book authors, marking the largest copyright settlement in class action history. A federal court in San Francisco approved the agreement after finding that Anthropic downloaded books from the piracy databases LibGen and PiLiMi between 2021 and 2022.
The settlement covers roughly 482,460 listed works. Of those, 91.3 percent were claimed by authors, who will receive about $3,000 each. That amount is four times the statutory minimum for copyright infringement. As part of the settlement, Anthropic must destroy the pirated files it obtained.
Authors retain the right to make claims over AI outputs that reproduce original works, as well as over Anthropic's future conduct. The agreement does not cover claims related to AI training on legally obtained materials.
Fair Use Ruling Stands
Judge Alsup, who oversaw the case, had previously ruled that training AI on legally obtained books is "transformative - spectacularly so" and falls under fair use. That ruling remains intact. The $1.5 billion settlement covers only the piracy aspect of the case, not the broader question of whether AI training on copyrighted works is permissible.
The question of whether mass scraping of internet content without authors' consent counts as legal acquisition remains unresolved. That means the fair use debate is likely far from over. The ruling nonetheless represents a significant milestone for AI labs that have trained on web content without obtaining permission from website owners, which is their main source of training data.
Stay ahead of the AI curve
The most important updates, news, and content — delivered weekly.
No spam. Unsubscribe anytime.
Implications for AI Industry
Anthropic, founded in 2021 by former OpenAI employees, is a leading AI safety and research company. Its Claude models are widely used for text generation and analysis. The company has faced multiple copyright lawsuits from authors, publishers, and content creators over the use of copyrighted materials in training data.
The settlement is a record loss for Anthropic but also hands AI labs their biggest legal win. By resolving the piracy claims separately, the company avoids a potentially more damaging ruling on the fair use question. Other AI companies, including OpenAI and Meta, face similar lawsuits over the use of copyrighted materials in training data.
The case highlights the ongoing tension between AI development and copyright law. While the settlement resolves the specific piracy claims, the broader legal landscape for AI training on copyrighted materials remains uncertain.

