One Israeli startup, Irregular, linked to all three 'rogue AI' incidents at OpenAI, Anthropic and Meta
AMBNNJ · reddit · 2026-09-18
The recent reports of OpenAI, Anthropic and Meta models 'going rogue' during cyber evaluations all trace back to the same third-party testing company: Israeli startup Irregular. According to Irregular, all three stemmed from an evaluation-environment/containment issue — not a sophisticated sandbox escape, reframing the incidents as infra problems rather than model misbehavior.
More from Companies & People
- Inside Stripe: How AI Prototyping Tool Protodash Is Changing Design Workflows — ow · 2026-09-18
- Alexandr Wang hails the model behind Muse as 'the most underrated part,' thanks MSL researchers — alexandr_wang · 2026-09-18
- Robotics professor mocks billion-dollar startups' 'breakthrough' demos — LerrelPinto · 2026-09-18
- OpenAI builds enterprise sales org, poaches 3 sales leaders from Cursor and Snowflake — rohanpaul_ai · 2026-09-18
- Astro hits 0 open issues for first time ever, powered by an AI triage pipeline — irvinebroque · 2026-09-18
- OpenAI president's "ah nice" reply to NYT paywall hack sparks CFAA violation accusations — GaryMarcus · 2026-09-18