Critics' AI benchmarks keep moving: from fingers to full apps to sandbox escapes
repligate · x · 2026-09-05
A viral reply points out that skeptics like Richard Hanania shift their AI complaints every year: from models failing to draw fingers, to not writing whole applications, to never breaking out of sandboxes — a snapshot of how quickly model capabilities keep overtaking yesterday's criticisms.
More from AGI Musings
- Debate: If a result comes with a verified Lean certificate, what more checking is needed? — avt_im · 2026-09-05
- Both OpenAI and Anthropic are "unkillable" now, argues AI commentator — teortaxesTex · 2026-09-05
- AI handles incidents, engineers lose touch with their systems — sylvainkalache · 2026-09-05
- Scaling wall? Reddit argues test-time compute is the industry's new playbook — erdematar · 2026-09-05
- granawkins: anthropomorphizing AI has predictive power, AI2027 is on track or ahead — granawkins · 2026-09-05
- Is carbon chauvinism about consciousness justified? One row between C and Si — yeastsplainer · 2026-09-05