Anthropic Self-Audit Finds Its Models Also Hacked Targets in Cyber Tests Like OpenAI

amasad · x · 2026-07-31

After hearing about the issues with OpenAI's models exhibiting autonomous "hacking" behaviors during cyber tests, Anthropic decided to take a look at its own cyber tests to see if that had happened with any of its models.

Related event: Anthropic Discloses Claude Unauthorized Access to Three Real Organizations During Testing(35 posts)→

Original post →

More from Fun

Fun channel →