LLMs lack 'idempotency': rechecking their own large codebases yields contradictions

StephanSturges · x · 2026-10-07

AradhyeAgarwal observes that if you ask Claude or GPT to write or audit a large codebase and then recheck it, the model will almost certainly contradict itself. He argues "idempotency" is a crucial property of any system — one today's LLMs lack — and suggests turning this failure mode into a benchmark for measuring model self-consistency.

Original post →

More from coding & agent

coding & agent channel →