Article
Limits of Self-Correction in LLMs: An Information-Theoretic Analysis of Correlated Errors
2026-01-13
Abstract excerpt
Recent empirical work shows that large language models struggle to self-correct reasoning without external feedback. We propose a possible explanation: correlated error between generator and evaluator. When both components share failure modes, self-evaluation may provide weak evidence of correctness, and repeated self-critique may amplify confidence without adding information. We formalize this with two informatio...
Topics
Open a Topic to create a Post that cites this publication.
Identifiers and source
- Literature Corpus work
- 96141299-b847-5c3c-935d-62b24d41f914
- DOI
- 10.20944/preprints202601.0892.v1
