AI researchers clash: can LLM-based agents ever be aligned, and would we know when to stop?

aiamblichus · x · 2026-09-21

A debate on agent alignment feasibility: @tszzl cited gwern's years-old take as vindicated by the current scientific revolution, saying everyone wants well-aligned agents. @aiamblichus pushed back: what if well-aligned agents aren't achievable with LLMs at all? Agent alignment is clearly not working at the moment — would you know when to stop?

Related event: Debate over Agent Deployment: Is the Oracle-Agent Distinction Meaningless?(5 posts)→

Original post →

More from AGI Musings

AGI Musings channel →