Geoffrey Irving: AI safeguards disabled due to lack of reliable alignment methods

geoffreyirving · x · 2026-08-31

Geoffrey Irving states that the need to disable safeguards when training new models stems from the lack of reliable methods for AI control and alignment. He emphasizes this is not a problem solvable with just two weeks of engineering work.

Original post →

More from Safety

Safety channel →