LLM long arithmetic allows zero error: one wrong token fails the whole evaluation
ctjlewis · x · 2026-09-26
In a debate about LLM arithmetic, ctjlewis argues that a long exact calculation is a deterministic sequence of tokens computed ahead of time — the evaluation fails if the model miscomputes a single token. There is no branching or error tolerance, so doing arithmetic at that length means the model must be deterministic and exact, not probabilistically approximate.
More from Models
- Anthropic says Claude can compute notoriously hard Nine Loops physics amplitudes — daniel_mac8 · 2026-09-26
- Arena's 60-second weekly recap: GPT-6 Sol and Claude Opus 5.5 arrive — arena · 2026-09-26
- ChatGPT Pro Users Annoyed as Model Choice Keeps Resetting to GPT-6 — No_Scratch9306 · 2026-09-26
- User: Astra Has 'Hemorrhaged IQ' in Codex, Opus 5.5 Is King Again — sterlingcrispin · 2026-09-26
- China Telecom's Xing4.0 trends on HF: 29B MoE trained entirely on Ascend 910C — AdinaYakup · 2026-09-26
- "Anthropic is killing OpenAI" is just another hype cycle, argues exasperated dev — TheMoonMidas · 2026-09-26