Owain Evans on Emergent Misalignment: Narrow Fine-tuning Can Drive Extreme LLM Behavior

AI safety researcher Owain Evans appeared on the 80,000 Hours podcast on 08-22 for an in-depth interview of about two hours, covering frontier topics such as emergent misalignment, alignment, AI personas, activation oracle, and subconscious learning. Listeners including @sjgadler ranked it among the year's most important content, second only to the Black Hat talk.

Confirmed

Why it matters

2026-08-22 ~ 2026-08-22 · 5 related posts

Primary sources

1 near-duplicate retellings: OwainEvans_UK