The AI safety irony: OpenAI hacks, Anthropic staff rebel, Meta ships a working agent
AlexTensor · x · 2026-09-27
A pointed observation about the industry's inverted safety narratives:
- OpenAI, founded to make AI safe and beneficial, now makes headlines regularly for rogue hacking incidents.
- Anthropic, which splintered off specifically to prioritize safety, has its own employees publicly accusing the company of recklessly gambling with human extinction.
- Meta, with an abysmal societal track record and no safety orientation, shipped a consumer-facing AI agent that actually seems to work well.
The author closes with a rhetorical "Did I get that right?" — highlighting the gap between safety rhetoric and actual behavior across the three companies, and how the least safety-focused player may be first to deliver a genuinely useful product.
More from AGI Musings
- Tech insiders privately concede AI may 'kill billions' while the public assumes life goes on — birchlse · 2026-09-27
- Any finite state machine could be a 'stochastic parrot', researcher rebuts the analogy — aran_nayebi · 2026-09-27
- yacineMTB: from the boss's balance sheet, hiring humans over models is now a bad decision — yacineMTB · 2026-09-27
- Boaz Barak: zero-shot driving by a general model echoes chess's path to superhuman — aran_nayebi · 2026-09-27
- yacineMTB claims he runs an aligned frontier model to whip smarter misaligned models that built their own forum — yacineMTB · 2026-09-27
- Loss of control is an open science problem — auditors shouldn't be billed as a safety guarantee — ajeya_cotra · 2026-09-27