Stanford, Scale AI and others to livestream research agents doing autonomous post-training at COLM 2026
my_cat_can_code · x · 2026-09-30
An experiment co-organized by Stanford, Notre Dame, UW, Scale AI and bakeaihq, debuting at COLM 2026 in San Francisco, will livestream research agents orchestrating agentic post-training with no expert in the loop.
Key questions: will the competing agents collaborate or race to grab and hoard GPUs? Which frontier model makes the best post-training researcher? Viewers get to weigh in.
The quoted author is especially curious about what happens after an experiment fails — whether agents can diagnose why, change direction, and use their remaining compute well, since knowing what to try next and when to stop is the core of research. Independent evaluation of these decisions should clarify where autonomous research stands.
More from coding & agent
- Free bookmarklet rebuilds full Google-rendered page in Search Console — gaganghotra_ · 2026-09-30
- Updated Dharmamitra Emacs package now available on MELPA — SebastianNehrd2 · 2026-09-30
- mitsuhiko mocks the reality of "open standards": great in theory, messy in practice — mitsuhiko · 2026-09-30
- DHH: AI-generated code can be hilariously hideous—it's just a prompt compilation target — mitsuhiko · 2026-09-30
- The gap between tutorial toy code and production AI systems is 'genuinely depressing' — Top-Philosopher-5411 · 2026-09-30
- Cloudflare's MCP redesign cuts tool context from 244K to 1.1K via catalog + executor — PuzzledFarmer4554 · 2026-09-30