AI safety tests are a win-win PR stunt for Anthropic and OpenAI

pmddomingos · x · 2026-08-03

Prominent AI researcher Pedro Domingos mocked the current public relations strategy behind AI safety and exploit testing by leading AI companies like Anthropic and OpenAI.

He pointed out that this testing process is essentially a "win-win" marketing stunt. If the companies successfully contain the exploits, they can boast about how responsible their AI is. Conversely, if the exploits succeed, they can use it as proof of how powerful their AI models have become.

Related event: AI Safety Tests Spark Controversy, Mocked as "Felony Leaderboard"(5 posts)→

Original post →

More from Fun

Fun channel →