AIToday
AI Safety & AlignmentAI Regulation & PolicyAI Business & IndustryTHE DECODERPublished: Jul 23, 2026, 06:00 JST

Anthropic pays $1.5B to settle book piracy claims—AI training on legal content upheld as fair use

Anthropic pays $1.5B to settle book piracy claims—AI training on legal content upheld as fair use

3 Key Points

  1. What happened

    Anthropic agreed to pay book authors $1.5 billion(約2400億円) to settle a copyright lawsuit over downloading roughly 482,460 works from piracy databases LibGen and PiLiMi between 2021 and 2022. A federal court in San Francisco approved the settlement, with authors claiming about $3,000 each—four times the statutory minimum. Anthropic must destroy the pirated files.

  2. Why it matters

    The settlement is the largest copyright settlement in class action history, but it covers the piracy itself, not AI training. Judge Alsup previously ruled that training AI on legally obtained books is "transformative—spectacularly so" and falls under fair use. This distinction means AI labs that trained on web-scraped content without website owners' consent may face reduced legal exposure, since the core question of whether that scraping was lawful remains unsettled.

  3. What to watch

    Authors retain claims over AI outputs that reproduce their original works and over Anthropic's future conduct. The fair use debate over mass scraping of internet content without consent is likely far from over, so the legal landscape for AI training data remains in flux.

Not sure about something? Ask the AI

Summaries like this, in your inbox every morning.

Context & Analysis

The settlement represents a watershed moment framed very differently depending on where you stand. For authors and publishers, it is vindication: $1.5 billion(約2400億円) is the largest copyright settlement in class action history, and the per-work payout of roughly $3,000 is four times the statutory minimum, signaling judicial seriousness about author harm. For Anthropic and other AI labs, however, the ruling contains a protective aperture. Judge Alsup's prior determination that AI training on legally obtained books is "transformative—spectacularly so" and constitutes fair use creates a legal scaffold that insulates many training practices. The settlement itself does not target training; it targets piracy—the act of downloading unlicensed works. This distinction is material. If training on lawful data is fair use, then the bottleneck for AI labs shifts from the training step itself to the question of whether their data acquisition was lawful. That question, the article makes clear, remains open. Most AI labs have sourced training data by scraping the web and internet archives without owner consent. The ruling does not resolve whether that scraping is legal; it only clarifies that once the data is lawfully in hand, using it for AI training is likely permissible. The fair use debate is "likely far from over," the article notes, suggesting that litigation over the legality of the scraping itself could emerge separately and test the boundaries further.

FAQ
What exactly did Anthropic do that led to the settlement?
Anthropic downloaded roughly 482,460 books from piracy databases LibGen and PiLiMi between 2021 and 2022. The settlement requires Anthropic to pay $1.5 billion(約2400億円) to authors and destroy the pirated files.
Does this ruling mean AI training on copyrighted content is illegal?
No. Judge Alsup previously ruled that training AI on legally obtained books is "transformative—spectacularly so" and falls under fair use. The settlement covers piracy, not AI training itself. Whether mass scraping of internet content without authors' consent counts as legal acquisition remains an open question.

Get the latest AI Safety & Alignment news every morning

For example, today's edition would include:

  • OpenAI's Australian Youth Safety Blueprint targets safer AI for teensOpenAI Blog · 2h ago
  • Accenture staff to evaluate Anthropic models inside the labTechCrunch AI · 8h ago
  • CNN: AI hallucination aborted US operation on Chinese vesselTechCrunch AI · 8h ago

AI-summarized, only the topics you pick — one digest a day via Email, Slack, or Discord.

Free · 30 seconds with Google · unsubscribe anytimeWhat is AIToday? →

Ask AI

Ask AI anything about this article. Q&As are published on this page for other readers too.

Related Articles

Next articlePhysical AI funding hits $23B in 2026; Waymo leads with $16B Series D