Alignment researcher's 2022 "moment of terror": GPT-4 done training, mentors quit OpenAI

JacquesThibs · x · 2026-09-14

AI safety researcher JacquesThibs recounts his August 2022 "moment of terror": days of backcasting alignment on a whiteboard kept pointing to deception and Goodharting derailing everything; his mentor, spooked right after GPT-4 finished training, quit OpenAI believing shorter timelines left nothing to do from inside; another former mentor, aware of the completed training and an upcoming capability jump, became convinced of a 2026 timeline. Notably confirms GPT-4 had finished training months before release.

Original post →

More from AGI Musings

AGI Musings channel →