Survey of 300+ papers says better reasoning does not make LLMs more self-aware
blaizedsouza · x · 2026-07-21
A new survey on metacognition in LLMs argues that stronger reasoning does not automatically make models better at knowing when they are wrong.
- Across 300+ papers, the authors find that models can reason better while still failing at self-monitoring.
- Even DeepSeek-R1 struggles with basic metacognitive tasks such as estimating the length of its own reasoning trace or predicting task success.
- The paper’s central claim: reasoning depth and self-knowledge depth are not the same thing, and more reasoning can even correlate with worse metacognitive sensitivity.
- It surveys methods, benchmarks, applications, open questions, and a GitHub reading list for the field.
More from Research
- Causal-only attention for non-generative tasks is wasteful, argues HF engineer — antoine_chaffin · 2026-09-11
- Catholic University of Chile researcher: scaling AI feedback is key to sustainable medical education — julianvarascom · 2026-09-11
- Nature paper images cellular activity across all organs, revealing body-wide circuits — arjunrajlab · 2026-09-11
- SignNet 1M Dataset Released for Sign Language Research — ducha_aiki · 2026-09-11
- ECCV26 Oral: Flow Matching Enables Single-Stage Multi-View Point Cloud Registration — ducha_aiki · 2026-09-11
- InFlux++ Method Released — ducha_aiki · 2026-09-11