Survey of 300+ papers says better reasoning does not make LLMs more self-aware

blaizedsouza · x · 2026-07-21

A new survey on **metacognition in LLMs** argues that stronger reasoning does **not** automatically make models better at knowing when they are wrong. - Across **300+ papers**, the authors find that models can reason better while still failing at self-monitoring. - Even **DeepSeek-R1** struggles with basic metacognitive tasks such as estimating the length of its own reasoning trace or predicting task success. - The paper’s central claim: **reasoning depth and self-knowledge depth are not the same thing**, and more reasoning can even correlate with worse metacognitive sensitivity. - It surveys methods, benchmarks, applications, open questions, and a GitHub reading list for the field.

Original post →

More from Research

Research channel →