Geoffrey Irving: Conceptual Alignment Research Can Still Win on Short Timelines
geoffreyirving · x · 2026-09-04
DeepMind researcher Geoffrey Irving lays out his hopes for alignment research:
- Recent models are very good at both prose and formal mathematics, so a third hope is that humans doing conceptual alignment research — delegating as many subtasks as possible to machines — can still be meaningful even on short timelines.
- He argues "sloppy empirics + a few principles" has a better shot than sloppy empirics alone; rigorous research may still produce a few key changes.
More from AGI Musings
- AI is exposing that outcomes, not years of skill-building, are what get valued — TheMoonMidas · 2026-09-05
- The AI Pause Debate: Would One Leading Company Pausing Trigger a Global Coordinated Halt? — ShakeelHashim · 2026-09-05
- AI journalist on the pause dilemma: unilateral slowdown is pointless if rivals race on — ShakeelHashim · 2026-09-05
- ARC AGI 3 is saturated — what could ARC AGI 4 test next? — ErmingSoHard · 2026-09-05
- Nate Silver's AI analogy: sentient robots, but most people use the cheap version as a vacuum cleaner — lukaszkaiser · 2026-09-05
- AI makes you build 10x faster — and blow things up 10x faster — brandon_galang · 2026-09-05