Luiza Jarovsky: Beyond a capability threshold, AI alignment is likely impossible

LuizaJarovsky · x · 2026-10-07

AI researcher Luiza Jarovsky shares a controversial take: after a certain capability threshold, AI alignment is likely no longer possible. The implication is that alignment work must succeed before models reach that point, since aligning super-capable systems afterward may be infeasible. The post is brief and doesn't lay out detailed arguments.

Original post →

More from AGI Musings

AGI Musings channel →