New CET method traces how psychological constructs emerge across LLM layers
GolinoHudson · x · 2026-09-12
A new interpretability method, Construct Emergence Tracing (CET), combines dynamic exploratory graph analysis, NMI, and network complexity measures to track how multidimensional psychological constructs organize across LLM layers. Across 14 checkpoints (117M–32B) and 10.6M network estimates, construct recovery generally rose from early to intermediate states — but five models (GPT-2, Phi-4-mini, Qwen2.5-32B, GPT-OSS-20B, Muse-Glimmer-30B) showed near-zero final-state NMI, showing parameter count and last-layer extraction alone give incomplete pictures.
More from Research
- Clay Math Institute says the Navier-Stokes problem 'has apparently been settled' — badumtsssst · 2026-09-12
- Decagon shares 19+ ablations on using GEPA for test-driven prompt optimization in production — kastnerkyle · 2026-09-12
- GraphED: graph-based AI learns how solids deform by sharing law structure across materials — bravo_abad · 2026-09-12
- GeoGuessr as an RL env: 4B VLM trained with OpenEnv and TRL to play the game — SergioPaniego · 2026-09-12
- Blur-to-video: SIGGRAPH Asia 2025 work recovers past, present and future frames from one motion-blurred photo — CSProfKGD · 2026-09-12
- IEEE Spectrum revisits how Lotfi Zadeh defied his critics to invent fuzzy logic — ArtificialOther · 2026-09-12