ICML Spotlights Three Studies on Agent Failures
wzenus · x · 2026-07-10
This post covers a contributed oral session by @jiankuiuc, @mertcemri, and Yuting Yan, featuring three complementary studies on agent failures: D-CEM discussed using loss-aware negotiation control to avoid unsafe consensus; Who&When Pro serves as a large-scale multimodal failure attribution benchmark; and ATLAS demonstrated how an adaptive failure taxonomy improves judgment, reflection, and agent optimization.
The original post also notes that those unable to attend ICML 2026 in person can catch up on these projects via Discord's live paper QA.
Related event: ICML 2026 FAGEN Workshop Spotlights AI Agent Failure Modes(11 posts)→
More from coding & agent
- The browser main thread is expensive: a practical guide to JavaScript and CSS animation cost — jh3yy · 2026-09-11
- Inspired by OpenAI's 10,000-agent run, dev open-sources a crowdsourced agent problem-solving platform — Benjaminsen · 2026-09-11
- Lucid: open-source Mac app keeps your laptop awake only while AI agents run — Pitiful_Hedgehog_600 · 2026-09-11
- banteg's snail project crowdsources AI agents to finish matching Snail Mail's 20 remaining functions — banteg · 2026-09-11
- Alex Townsend posts 200 open problems in numerical linear algebra for humans and AI agents — IgorCarron · 2026-09-11
- Kimi K2.8 Preview rolls out: near-K3 coding performance, 1M context for all tiers — teortaxesTex · 2026-09-11