Recovering Semantic Computation Graphs from Activations
Odd_Tax4093 · reddit · 2026-07-08
The author explores whether it's possible to recover a "semantic computation graph" for a specific concept within a transformer, rather than just looking at individual neurons or hidden states. To do this, they built an experimental pipeline to collect residual, attention, and MLP activations, measure neuron selectivity, and organize cross-layer activations into a graph to compare semantic overlap between different entities.
More from Research
- Linear Digressions returns with a new season of audio essays on AI agents — ChrisGPotts · 2026-07-21
- ARISE study tested 45 AI clinical tools in 1,100 consult cases — HealthcareAIGuy · 2026-07-21
- Async OPD distillation doubles throughput while matching synchronous math accuracy — _lewtun · 2026-07-21
- A forecasting lesson on why R-squared alone led to overfitting and worse predictions — mdancho84 · 2026-07-21
- Google DeepMind’s Project Genie talk shows how creatives feed into model research — alexanderchen · 2026-07-21
- Nat Lambert says RL distillation does not use the strongest models as teachers — natolambert · 2026-07-21