A new LLM paper learns manifold features instead of linear SAE directions

Sauers_ · x · 2026-07-27

The post summarizes a paper on learning a manifold dictionary in LLMs without supervision.

The result is framed as a more expressive alternative to linear SAE steering, especially for cyclic or non-linear concepts like weekdays.

Original post →

More from Research

Research channel →