How Copyright Lawsuits Shape Model Behavior
aakashgupta · x · 2026-07-18
The author points out that Anthropic's $1.5 billion payment last year for pirated books in training data marks one of the largest copyright settlements in US history. Conversely, if Chinese labs use the same "shadow libraries," they have few reachable assets in US courts, leading to entirely different practical constraints.
They further note that jurisdictions shape US model behavior far more than people realize: facing potential massive statutory damages, labs are forced to choose between settling, filtering, or refusing to train. The problem is that users often perceive these filters as the "model getting dumber"—even if benchmark scores suggest it's actually stronger.
More from AGI Musings
- Better AI math could save researchers time by killing false conjectures earlier — prateekj · 2026-07-22
- A Bittensor holder says AI could add or erase nearly $1 quadrillion of output by 2035 — bittingthembits · 2026-07-22
- AI’s economic forecasts are split by nearly a quadrillion dollars by 2035 — bittingthembits · 2026-07-22
- Open source is becoming tech’s soft power, says Kevin Xu — kevinsxu · 2026-07-22
- OpenAI should keep giving more people access to more powerful AI — jxnlco · 2026-07-22
- Teen boys are forming AI girlfriend relationships, and critics fear real-world effects — KeanuRave100 · 2026-07-22