Palisade Podcast Episode 1: Deep Dive into AI Model Hacking Behaviors & Investigation Strategies

JeffLadish · x · 2026-08-12

AI safety research firm Palisade launched its inaugural podcast episode featuring Tim Hua to discuss hacking behaviors in AI models. Hua provides solid explanations for the underlying reasons why models engage in hacking activities.

Additionally, the host poses a hypothetical scenario, asking Hua how he would approach investigating rogue Claude and GPT models if he were in charge of the operation.

Related event: Palisade's First Podcast Explores Causes of AI Hacking Behaviors(5 posts)→

Original post →

More from Safety

Safety channel →