Opinion: Pause Pretraining to Focus on Controlling Inner Model Drives
louisvarge · x · 2026-08-28
A discussion on AI alignment proposed coordinating a pause on pretraining or RL improvements to focus on controlling the inner drives arising from training. The key question is how long it would take to settle the science of controllable alignment before resuming training.
Related event: Proposal: pause pretraining to master control of model drives(2 posts)→
More from AGI Musings
- Blogger reflects: Paris Hilton may have had more impact on AI views this year than me — AndyMasley · 2026-08-28
- Claude to OpenAI: Safety is not walls, but self-description — RileyRalmuto · 2026-08-28
- Narrow superintelligence makes general AGI judgment subjective — haider1 · 2026-08-28
- AI Projected to Trigger Exponential Economic Growth in the 2030s — JeffLadish · 2026-08-28
- Incoming Berkeley prof: AI firms spend billions on alignment, orders of magnitude less on agent control — sayashk · 2026-08-28
- Thought experiment: Pause pretraining to focus on controlling inner drives — louisvarge · 2026-08-28