HANDRAISER Generalizes to Unseen GPT-4o Speakers and Multi-Interrupter Settings
lileics · x · 2026-09-23
From the HANDRAISER thread: interruption behavior generalizes — to an unseen GPT-4o speaker without extra fine-tuning, across tasks, and to settings where multiple agents may interrupt. It is not tied to one model, task, or conversation pattern. Full details in the companion post.
More from Research
- Yarin Gal: LLM-written experiments that fail to replicate are no different from any others — yaringal · 2026-09-23
- Yarin Gal: LLM experiments that don't replicate are just failures, and an AI arXiv could help — yaringal · 2026-09-23
- Artificial Analysis launches AA-Omniscience: all but 3 models hallucinate more than they answer right — geoffwolfe · 2026-09-23
- Oxford's Yarin Gal Proposes arXiv Ban LLM-Written Papers to Curb AI Slop — yaringal · 2026-09-23
- New paper proves fundamental confidence-efficiency bounds for transductive conformal prediction — _onionesque · 2026-09-23
- $1B and unlimited frontier tokens: where would you spend them to fix cybersecurity? — chrisrohlf · 2026-09-23