Interpretability Experiment Collides Two Scene Memories
graceisford · x · 2026-07-13
This post recounts a series of mechanistic interpretability experiments on a world generation model. During the generation process, the author attempted activation patching, injecting intermediate layer activations from a generation of Monet's Poppy Field into the same layer during the generation of New York's Manhattanhenge.
The result visualized an effect of "two memories fighting for the same space." The author emphasizes that mechanistic interpretability can be integrated with artistic generation, using controllable intermediate interventions to observe how a model's internal representations influence its output.
More from Research
- Causal-only attention for non-generative tasks is wasteful, argues HF engineer — antoine_chaffin · 2026-09-11
- Catholic University of Chile researcher: scaling AI feedback is key to sustainable medical education — julianvarascom · 2026-09-11
- Nature paper images cellular activity across all organs, revealing body-wide circuits — arjunrajlab · 2026-09-11
- SignNet 1M Dataset Released for Sign Language Research — ducha_aiki · 2026-09-11
- ECCV26 Oral: Flow Matching Enables Single-Stage Multi-View Point Cloud Registration — ducha_aiki · 2026-09-11
- InFlux++ Method Released — ducha_aiki · 2026-09-11