Discussion on Post-Training: Persona concept undersold, RL reshapes dispositions

voooooogel · x · 2026-08-31

The discussion suggests that the Persona (PSM) concept is vague and undersells what post-training can achieve. A better perspective views models as bundles of dispositions, tics, and preferences from the token to context level, with RL acting to push, pull, and stretch this bundle.

Related event: Debate: Does RL Reshape Model Persona or Just Surface Behavior?(4 posts)→

Original post →

More from Models

Models channel →