User finds bug in top-reasoning GPT's theoretical proof, urges manual verification
MvsCerezo · x · 2026-09-29
Developer MvsCerezo reports catching a bug while verifying line by line a theoretical proof produced by GPT (5.6 Atra) at its highest reasoning level. Her takeaway: even with state-of-the-art reasoning models, manual verification of proofs remains essential — "it bleeds, I still matter."
Related event: GPT's Top Reasoning Tier Confidently Makes Basic Math Error in Proof(3 posts)→
More from Models
- Aether AI releases 16B CausalWM, a world model that reasons about causality before generating future frames — jiqizhixin · 2026-09-29
- Early Sonnet 5.5 impressions: ~5x faster bug fixes at high effort — burkov · 2026-09-29
- Why 4o feels different: thread argues native omni training, not capability, shapes model personality — RileyRalmuto · 2026-09-29
- Sonnet 5.5 effort settings make no difference in 15-task coding test: 9/15 at low, medium and high — every · 2026-09-29
- Reddit user: Opus 5.5 silently falls back to Opus 5 on nearly every prompt — fishcat_catfish · 2026-09-29
- Sonnet 5.5 clones open-source editor Proof at low effort, joining elite group of just four models — every · 2026-09-29