Personal Visual Reasoning Test Shows AI Far From AGI
adam_dorr · x · 2026-07-10
Developer Adam Dorr possesses a personal benchmark test for frontier AI models that no model has passed yet, including GPT 5.5, Opus 4.8, and Grok 4.5. The test features no related training data, making it an out-of-distribution problem that an average 12-year-old can easily solve. This indicates that current AI systems are still quite far from achieving true AGI.
More from AGI Musings
- ControlAI CEO says an international ban on superintelligence is needed to avert extinction risk — zetalyrae · 2026-07-22
- Gary Marcus says LLMs still cannot really do math on their own — GaryMarcus · 2026-07-22
- Gary Marcus says LLM math skills are like knowing only a car’s engine size — GaryMarcus · 2026-07-22
- AI may make digital work infinitely leveraged while offline life gets more human — illscience · 2026-07-22
- Better AI math could save researchers time by killing false conjectures earlier — prateekj · 2026-07-22
- AI’s economic forecasts are split by nearly a quadrillion dollars by 2035 — bittingthembits · 2026-07-22