Highlights from the Agent Safety Workshop
wzenus · x · 2026-07-10
The post recaps a presentation from the ICML 2026 Failure Modes in Agentic AI (FAGEN) Workshop. The talk focused on safety in the agent era, mapping out the risk surfaces associated with agent memory, tools, private data, and real-world actions. It also introduced a safety pipeline featuring DecodingTrust, DTap adaptive red-teaming, runtime guardrails, and safety certifications.
Related event: ICML 2026 FAGEN Workshop Spotlights AI Agent Failure Modes(11 posts)→
More from Safety
- Anthropic publishes its most detailed threat report, including an AI-designed drone swarm case — soumitrashukla9 · 2026-09-11
- OpenAI asks Congress whether an industry-wide AI slowdown would be legal — The Decoder · 2026-09-11
- Author retracts 'a16z partner calls for nationalising frontier AI' post: likely a troll — S_OhEigeartaigh · 2026-09-11
- Houthis tried to use Claude to design missile software, Anthropic says it blocked the attempts — Affectionate_Bee6434 · 2026-09-11
- AI safety community mocked as 'bridge engineers' who say bridges can never be safe — Dan_Jeffries1 · 2026-09-11
- Why So Many AI Researchers Think the Machines Could Kill Everyone — wiredmagazine · 2026-09-11