LOCUS: targeted activation steering via head selection and subspace projection

FrancescoLocat8 · x · 2026-10-09

LOCUS addresses how to steer model behavior without degrading performance. It uses token subspaces tied to the target property to select which attention heads to steer and which subspace within each head, enabling precise, targeted activation steering.

Original post →

More from Research

Research channel →