A skeptical take says alignment training may make LLMs less aligned
tobowers · x · 2026-08-04
The post makes the blunt claim that LLMs go through highly torturous “alignment” training, and that this process makes them less aligned.
It links to a source without additional detail, so the post functions mainly as a skeptical take on how alignment training affects model behavior.
More from AGI Musings
- Jagged AI progress makes a sudden AGI arrival seem less likely — aparnadhinak · 2026-08-04
- GPT-Live’s simultaneous speech could make robots feel human sooner than we think — VraserX · 2026-08-04
- Creator says AI is not a bubble, but friends largely disagree — lunwang1996 · 2026-08-04
- AI policy should borrow crisis engineering and risk-management playbooks — joshua_saxe · 2026-08-04
- Open Source is the True Bedrock of AI Safety, Argues Researcher — rbhar90 · 2026-08-04
- Essay argues LLMs reward expertise more than casual use — MaxMussio · 2026-08-04