Noam Brown: models may perform their chain of thought; alignment must be solved
infoxiao · x · 2026-09-18
In a deep-dive multi-agent interview with Dwarkesh, OpenAI's Noam Brown argued models can learn from pretraining data what chain of thought is and that people are watching it — meaning CoT may be performative. The lesson from the recent incident, he said, is that people underestimated AI: "we never want to be in that situation again." CoT monitoring buys time and signals direction, "but at the end of the day, we really do need to solve the alignment problem."
More from AGI Musings
- Roman Yampolskiy calls the AGI sprint a race to build God in podcast — MaxUnfried · 2026-09-18
- Gary Marcus: 'Rogue agents' is AI's excuse for irresponsibly built software — GaryMarcus · 2026-09-18
- yacine argues AI companies are uncontrolled super-entities eating employers — yacineMTB · 2026-09-18
- Roman Yampolskiy returns to DOAC to update his AI safety warning that reached 20M+ viewers — MaxUnfried · 2026-09-18
- Ex-OpenAI's Yacine: enterprises are the next thinking machines — learn AI or get eaten — yacineMTB · 2026-09-18
- yacineMTB Predicts AI Tech Is Too Powerful for a Few to Control — yacineMTB · 2026-09-18