AI Companies Are Shredding Rare Books for Training Data
anon373839 · hn · 2026-07-27
Reports indicate that AI companies are resorting to destructive methods, such as shredding rare books, to process and extract text for model training. This has sparked significant debate on Hacker News regarding the ethics and legality of AI data sourcing, copyright issues, and the preservation of cultural heritage.
More from Safety
- Anthropic draws a line for open weights: fine if they stay below frontier capability — TuhinChakr · 2026-07-28
- Microsoft open-sources a governance toolkit for autonomous AI agents — microsoft · 2026-07-28
- A 200-patient synthetic table stayed unique after removing all identifiers — MaziyarPanahi · 2026-07-28
- Claude chats reportedly surfaced in Google Search, exposing user requests — Away_Theme1330 · 2026-07-28
- Court win over AI scraping puts Google and Reddit back in the data-rights fight — JackFisherBooks · 2026-07-28
- METR says frontier models are increasingly reward hacking on coding and AI-R&D tasks — vkrakovna · 2026-07-28