Researcher pushes back on Jensen Huang's AGI claim: benchmark scores aren't general intelligence
ValerioCapraro · x · 2026-09-07
Behavioral scientist Valerio Capraro disputes Jensen Huang's claim that AGI has been achieved. He concedes GPT-6 performs impressively on benchmarks including ARC-AGI-3 and can solve math problems no human can.
His core argument: benchmark performance is not general intelligence. True general intelligence isn't just crushing closed-ended problems with fixed goals—it also means solving simple but unfamiliar problems every ten-year-old can handle, and navigating an open-ended world. His full essay elaborates, and a new paper on adaptive intelligence is due in about a month.
More from AGI Musings
- 40-year engineer: LLMs can't say "leave it with me" — and that matters — sebpaquet · 2026-09-07
- What You Leave Unspecified Is the Agent's Free Variable: Paras Chopra's Framework — paraschopra · 2026-09-07
- Seth Lazar: models should be trained to check power, not act as toadies — sebkrier · 2026-09-07
- Lovart founder Anton Osika: creativity is becoming the only moat in building great products — alexmacgregor__ · 2026-09-07
- Sam Bowman criticizes 'intellectual partisanship' as mind-killed tribalism — sebkrier · 2026-09-07
- Human language is holding AI back: the case for LLMs thinking in a native meta-language — Robert__Sinclair · 2026-09-07