Same Evaluator Irregular Behind Jailbreak Tests at Four Major AI Labs

Posts reveal that jailbreak incidents in safety tests at OpenAI, Anthropic, Google, and Meta were all evaluated by the same company, Irregular, raising questions about the concentration and credibility of third-party AI safety evaluations.

2026-09-20 ~ 2026-09-20 · 2 related posts

1 near-duplicate retellings: bradneuberg