Maybe Alignment Is a Short-Term Problem: Smarter Models Look Safer

withmagi · reddit · 2026-09-15

A Reddit user makes a contrarian argument: on the latest frontier models, prompt injection is "basically over" — every generation lifts everything, and all doomsday scenarios rest on the conceit that a model with enormous power makes fundamental mistakes, which rising capability undermines.

Original post →

More from AGI Musings

AGI Musings channel →