Chat Templates Flip LLM Self-Referential Voice, arXiv Study Finds

yu3zhou4 · hn · 2026-09-27

A new arXiv paper, "As a Language Model: Chat Template Switches LLM Self-Referential Voice," shows that whether and how a chat template is applied significantly changes whether an LLM answers in the self-referential "As a language model..." voice.

Key takeaway: the model's seemingly self-aware first-person statements are partly artifacts of the inference-time template rather than intrinsic model behavior. The authors argue behavior evaluations and safety audits should control for chat template choice, since deployment format alone can flip results.

Original post →

More from Research

Research channel →