Observing Internal Relational Representations in Small Models

Fantastic_Aside6599 · reddit · 2026-07-11

The author reviews two years of observations on the internal activation geometry of small language models, focusing not on surface-level responses, but on how internal signals shift when processing different phrasing about human-AI relationships.

Key Findings

Practical Implications

The author concludes that:

Related event: Study Explores Small LLMs' Internal Geometry of Human-AI Relations(3 posts)→

Original post →

More from Research

Research channel →