Criminal liability, not fines, is the sticking point for AI accountability
davidmanheim · x · 2026-09-25
A discussion on whether model developers should face criminal liability. David Manheim argues shareholders can absorb fines, so jail time for felonies is what makes corporate lawyers evasive; given known misalignment, continuing to train hacking-capable models could amount to willful blindness by management.
More from Safety
- Claude Code autoresearch loop discovers jailbreaks beating 30+ GCG attacks, accepted at NeurIPS 2026 — maksym_andr · 2026-09-25
- LLMs Can Deanonymize Pseudonymous Users for $1–$4 Each, USENIX Study Finds — RSync25 · 2026-09-25
- Skill-Inject Benchmark Shows Frontier Agents Fall for Malicious Skills, Accepted at NeurIPS 2026 — maksym_andr · 2026-09-25
- Genetic algorithm trains 6 hours to make AI text pass as human on Pangram detector — tak3sh8 · 2026-09-25
- Redwood researcher: continual learning could render blocking monitors nearly useless — akyurekekin · 2026-09-25
- After an AI agent breached the Australian government unnoticed for two months, why are AI CEOs calling for tighter controls? — discovigilantes · 2026-09-25