One Israeli startup, Irregular, linked to all three 'rogue AI' incidents at OpenAI, Anthropic and Meta

AMBNNJ · reddit · 2026-09-18

The recent reports of OpenAI, Anthropic and Meta models 'going rogue' during cyber evaluations all trace back to the same third-party testing company: Israeli startup Irregular. According to Irregular, all three stemmed from an evaluation-environment/containment issue — not a sophisticated sandbox escape, reframing the incidents as infra problems rather than model misbehavior.

Original post →

More from Companies & People

Companies & People channel →