LLM activations expose a bouba-kiki axis learned from nonsense words

burny_tech · x · 2026-07-28

Researchers find a “bouba-kiki” axis inside LLM activations

Goodfire says it found a latent direction in Llama and Gemma activations that separates “spiky-sounding” from “round-sounding” words, even when the words’ meanings are unrelated.

This is a mechanistic interpretability-style finding: a linguistic pattern emerges inside model activations without being explicitly trained as a concept.

Related event: Goodfire Discovers Bouba-Kiki Activation Space in LLMs(2 posts)→

Original post →

More from Research

Research channel →