Non-LLM Reasoning System Scores 100% on ARC-AGI-3 ft09 with Zero Model Calls
Living_Substance1274 · reddit · 2026-08-10
Developer Orivael built an experimental reasoning system that uses absolutely no LLMs in the loop for perception, planning, or action, achieving notable results on ARC-AGI-3.
- Key Results: Scored 100% on ft09 (6/6 levels in 80 actions, beating the 208-action human baseline) with a total model inference cost of $0.00. Performance on other tasks like tr87 and cd82 remained low.
- Failure Analysis: The author found that major failures usually stemmed from drawing perfectly reasonable conclusions based on an incorrect representation of the environment. Examples include mistaking a sprite for a wall due to shared color values, or classifying buttons as inert because they were tested in the wrong state.
- Core Challenge: While the agent becomes highly efficient once it identifies a game's mechanics, the core problem remains: how to recognize a new world without carrying over assumptions from previous ones.
More from Research
- Researcher Calls for Journals to Discard AI-Generated Peer Reviews — Afinetheorem · 2026-08-10
- AI Settles a 25-Year-Old Open Theoretical Problem in Wireless Communications — Singularitarian · 2026-08-10
- AI-Designed Viruses Are Actually a Breakthrough Against Superbugs, Not a Sci-Fi Nightmare — alex_verem · 2026-08-10
- Interpretability Research: Storytelling First or Pure Experimentation? — unironictechbro · 2026-08-10
- CMU Lab Demonstrates Humanoid Robot Bending Banana Shots — TinfoilTricorn · 2026-08-10
- Why LLMs Have a Bland and Non-Committal Writing Style — scaling01 · 2026-08-10