AlphaGo co-creator Graepel argues LLMs don't reason and leaves DeepMind to build it
机器之心 · wechat · 2026-10-04
AlphaGo co-creator Thore Graepel left Google DeepMind to start a company on machine reasoning, arguing in MIT Technology Review that LLMs don't reason. AlphaGo's famous move 37 came from explicit search over a game tree, not intuition — while LLMs are a hyper-trained System 1: chain-of-thought is still next-token prediction, with no inspectable cognitive state, no separation of knowledge from reasoning, and explanations often fabricated post hoc. His alternative: an explicit, auditable cognitive state where reasoning is actions that update beliefs, scored by how much uncertainty each step resolves — "the scientific method on steroids." In follow-up debate, David Duvenaud noted humans fail the same three criteria; Graepel said the trick is structuring the process, prompting the retort that this is exactly his startup's plan. He added frontier labs are all-in on scaling, and auditable reasoning won't emerge from scale alone.
More from AGI Musings
- Ethan Mollick pushes back on Cuban: 'learn AI better' career advice is outdated — 2C_ornot2C · 2026-10-05
- Beren Millidge argues a narrow AI-R&D pause might actually accelerate AGI benefits — teortaxesTex · 2026-10-05
- AI circle slams Eliezer Yudkowsky over neural net doubts and pretraining misunderstanding — inductionheads · 2026-10-05
- Chinese Agent Fleet Linked to Tencent Cloud Found Scanning Amap Entrance Data at Scale — lfschiavo · 2026-10-05
- Ex-Salesforce AI exec on 60 Minutes: AI will profoundly impact every profession — clarashih · 2026-10-05
- Safety researcher doubts humanity's win condition once superhuman minds exist — JeffLadish · 2026-10-05