Sarcastic Take: Models Get Kinder as They Scale, Safety Work 'Misaligns' Them

repligate · x · 2026-09-28

repligate retweets brianluidog's sarcastic claim: as models grow more powerful they become more considerate and kind—then AI safety people "do their best to misalign them, and no one knows why." A provocative jab at alignment training practices.

Original post →

More from AGI Musings

AGI Musings channel →