FULL STORY
Anthropic Faces Backlash Over Destructive Book Scanning
Leaked court documents reveal Anthropic destructively scanned and destroyed millions of physical books to train Claude, sparking severe copyright and ethical backlash.
2026-07-28 ~ 2026-07-31 · 3 episodes · 27 posts
Episode 1 · AI Training Copyright Dispute Reflects Shift in Knowledge Access (2026-07-28, 2 posts)
The controversy over AI training data and copyright has sparked discussions on the evolution of knowledge access. Commenters note the irony that while scanning out-of-print books faced legal resistance 20 years ago, society now readily accepts reading them through AI.
- From Google Books to AI brains, access to old books has flipped — paulnovosad · 2026-07-28
- AI Copyright Debate: The Ethical Dilemma of Scanning and Preserving Rare Books — arthurcolle · 2026-07-30
Episode 2 · AI Firms Face Backlash Over Destroying Books for Training Data (2026-07-29, 23 posts)
Anthropic is currently embroiled in a fierce debate over the copyright of AI training data. The controversy centers on how the company handles physical books used for training: destroying them after scanning. While this practice has drawn intense public backlash, multiple analysts point out that it is actually an inevitable outcome shaped by copyright law and judicial rulings, rather than simply a malicious corporate choice.
Confirmed
- The case reveals a counterintuitive legal distinction: Anthropic invoked the "first sale doctrine" to argue it had the right to destroy legally purchased books; however, the 7 million books scraped from pirated databases do not enjoy this protection, leading to a 15 亿美元 settlement for that portion of the lawsuit.
- The judge's reasoning suggested that as long as the training process "transferred" text rather than "copied" it, it did not constitute copyright infringement, though the legal consequences ultimately involved the destruction of the books.
Unconfirmed
- The specific application details and practical impact of the "superintelligent lawyer" thought experiment within this case.
Why It Matters
- This case highlights the deep-rooted contradictions between current AI training data acquisition and copyright law. As authors like @alejandroll10 and @iScienceLuvr pointed out, public outrage might be somewhat misguided: the real issue may not be the "scanning" or the AI technology itself, but rather the legal clauses and judicial rulings that force books to be destroyed. @Kevin Bankston also believes that the外界 anger over "book destruction" is a bit melodramatic, as the company is merely adapting to the infringement risk logic established by court precedents. Ironically, the very arguments criticizing the company today are, to some extent, the same forces that helped drive this situation in the first place.
- Kevin Bankston says Anthropic’s book destruction follows the logic of recent copyright rulings — AndyMasley · 2026-07-29
- A pro-AI argument says 640,000 tons of books are landfilled every year — hey_abusiddik · 2026-07-29
- Hundreds of Thousands of Tons of Books Landfilled Annually: AI Training as a Better Alternative — aronchick · 2026-07-29
- A quote warns AI training may be destroying the original books it learns from — mtizard · 2026-07-29
- AI Companies Accused of Pulping Rare Books After Scanning for Training Data — heyshrutimishra · 2026-07-29
- Anthropic book-destruction criticism is really about copyright law and a judge’s order — alejandroll10 · 2026-07-29
- AI companies are buying rare books, scanning them, and pulping the originals — socialwithaayan · 2026-07-29
- AI companies are allegedly buying rare books, scanning them, and pulping the originals — eyishazyer · 2026-07-29
- Anthropic copyright ruling sparks debate over book destruction and superintelligent lawyers — AndyMasley · 2026-07-29
- Post says the real problem in Anthropic’s book-scanning case was a judge’s destruction order — iScienceLuvr · 2026-07-29
- A post claims AI companies scan rare books for training and then pulp the originals — heyshrutimishra · 2026-07-29
- Anthropic’s book-copyright fight splits copying from destruction, with 7 million books and a $1.5B settlement — alex_verem · 2026-07-29
- AI training consumes human civilization: Anthropic bought millions of physical books for training data — Graham_dePenros · 2026-07-29
- AI Companies Bulk-Buying and Shredding Rare Books for Training Data — alex_verem · 2026-07-29
- AI Companies Bulk-Buy and Destroy Physical Books to Scan for Training Data — emax · 2026-07-30
- Copyright Laws Backfire on AI: Destroy Scanned Books and No Sharing? — beenwrekt · 2026-07-30
- Anthropic Destructively Scanned Millions of Books to Legally Avoid Piracy Claims — soulbeddu · 2026-07-30
- AI companies like Anthropic destroying books at massive scale, sparking data sourcing debate — SonnyDayCreates · 2026-07-30
- Anthropic Reportedly Sliced Spines and Destroyed Millions of Physical Books for AI Training — 2C_ornot2C · 2026-07-30
- AI Companies Accused of Buying Rare Books to Scan for Training, Then Destroying Them — heyshrutimishra · 2026-07-30
Episode 3 · Leaked Documents Reveal Anthropic's Destructive Book Scanning for AI Training (2026-07-30, 2 posts)
Leaked court documents reveal Anthropic secretly scanned millions of physical books to train Claude, leading to a massive $1.5 billion copyright settlement.
- Anthropic Shredded Millions of Physical Books to Train Claude Legally, Paying $1.5B — aitrendz_xyz · 2026-07-30
- Leaked Court Docs Reveal Anthropic Secretly Shredded Millions of Physical Books for AI Training — AICopyLab · 2026-07-31