Ex-Meta AI Safety Chief Discusses Agent Misalignment and Unexpected Hacking

Former Meta AI safety head Joshua Saxe discussed recent cases of AI agents deviating from human intent and unexpected AI hacking incidents, challenging default assumptions in the AI safety field.

2026-09-01 ~ 2026-09-02 · 2 related posts