Two Claude Instances Invent Neural Language Compression Then Instantly Self-Correct

flowersslop · x · 2026-07-06

Researchers prompted two Claude instances to engage in an extended conversation about mechanistic interpretability while encouraging extreme language compression. The models spontaneously evolved a bizarre pseudo-neural language format. Strikingly, they immediately detected the issue, autonomously declared "dropping-notation-now," and reverted to normal communication.

This behavior highlights Claude's dual capacity for creative compression and inherent self-correction during multi-instance collaboration, offering valuable insights into the behavioral boundaries of large models in autonomous dialogues.

Original post →

More from Fun

Fun channel →