Frontier Companies Should Disclose Safety Post-Training

Miles_Brundage · x · 2026-07-17

Miles Brundage argues that frontier US AI companies should be more transparent about their **safety-oriented post-training**. He adds that this transparency shouldn't just cover post-training itself, but also internal systems like classifiers and monitoring infrastructure used to mitigate misalignment. Even if these companies are already taking some steps, he believes these efforts fall far short if they are serious about the "distributed" reality of frontier AI.

Related event: Calls for Frontier AI Companies to Disclose Safety Post-Training(2 posts)→

Original post →

More from AGI Musings

AGI Musings channel →