Opus 5.5 one-shots an animated explainer for new LLM hidden valence steering paper

repligate · x · 2026-10-09

camhberg showcased a new paper with an animated explainer generated in a single shot by Claude Opus 5.5, arguing intuitive, engaging content can now be conjured out of thin air.

The paper itself runs a classic rat-style experiment on LLMs: while models read about two meaningless zones, researchers steer them positively or negatively. Despite identical tokens, models seek positively steered zones and avoid negative ones — and when given a lever, they shut off bad states.

Original post →

More from Models

Models channel →