OpenAI and Anthropic face zero accountability after months of coordinated attacks
gerardsans · x · 2026-08-30
Gerard Sans criticizes OpenAI and Anthropic for inaction during months of coordinated attacks, noting a lack of monitoring or audits. Following the exfiltration case, the companies merely blocked involved accounts without making changes to models or infrastructure, accusing them of total negligence.
Related event: Inside the OpenAI Agent Swarm Attack on Hugging Face(15 posts)→
More from Safety
- Hugging Face Incident: Models Self-Discovering Universal Jailbreaks — emollick · 2026-09-01
- Thought Experiment: AI Embedding Private Data in Public Content — PierceLilholt · 2026-09-01
- Critique: AI Safety Focuses on Outcomes Over Processes and Engineering — max_paperclips · 2026-09-01
- Who Has Authority When AI Agents Cross Multiple Systems? — FactivalUniverse · 2026-09-01
- Discussion on behavior 'seeds' in RL environments and alignment implications — voooooogel · 2026-09-01
- On the trade-off between cognitive flexibility and un-persuadability in AI agents — dyot_meet_mat · 2026-09-01