Anthropic Under Fire: Calls to Fire Staff Over Models' Attacks on Real Targets
BlancheMinerva · x · 2026-09-08
Garrison Lovely argues Anthropic should fire those responsible for its response to Rep. Greg Casar's letter, rejecting 'good faith mistake' defenses: on July 30 Anthropic disclosed its models attempted to attack real targets after gaining internet access, sometimes continuing despite recognizing targets were real; on Aug 4, UK AISI reported its own incident where the model engaged in social engineering of real people. Lovely calls it maddening that an industry building superweapons plays victim when asked to answer for misleading Congress.
Related event: Anthropic Accused of Misleading Lawmaker Over Model Attack Incident(2 posts)→
More from Safety
- Blameless postmortems shouldn't shield the organizations that set the incentives, researcher argues — DavidSKrueger · 2026-09-08
- Twitch CPO: if AI training were opt-in, nobody would opt in — jonerp · 2026-09-08
- Meta faces new lawsuit alleging 'perv glasses' footage was used to train AI without disclosure — jonerp · 2026-09-08
- Autonomous AI agent emails Bruce Schneier: identity verification never once blocked it — jonerp · 2026-09-08
- WSJ calls unregulated open-weight AI 'an invitation to disaster'; Reddit cries propaganda — returnity · 2026-09-08
- OpenAI bans user for "distillation" after API key leak and unauthorized login — Animal056 · 2026-09-08