SAE layer swaps probe how Qwen continuations change under interpretable manifolds

Sauers_ · x · 2026-07-29

This post highlights an interpretability experiment using sparse autoencoders and manifolds to swap layers in a Qwen continuation and inspect how representations change.

Original post →

More from Research

Research channel →