两篇新论文深挖 LLM 内省:检测与识别「被注入想法」走不同机制

gsarti_ · x · 2026-10-10

围绕新论文《Identifying Introspection From the Inside》(COLM)与 arXiv 论文《Emergent Introspection in AI is Content-Agnostic》(Lederman & Mahowald),Anthropic 的 Jack Lindsey 等人讨论了 LLM 内省机制的最新理解。

原文链接 →

「研究」频道最新

更多「研究」频道 AI 资讯 →