Agent Harness Aces ARC-AGI-3
naval · x · 2026-07-16
The Impossible Research team introduced an agent harness capable of playing games, writing code, and reasoning "like a physicist."
The post cites [schema] scores: achieving 99% RHAE on the ARC-AGI-3 Public set using Opus 4.8 + Fable 5, and 95.35% using GPT-5.6 Sol. The author's central claim is that this harness organizes LLMs into workflows better suited for high-difficulty reasoning and task execution.
Related event: Schema Harness Sparks ARC-AGI-3 Debate(14 posts)→
More from coding & agent
- Claude Code skill uses 10 Markdown rules to make outputs ADHD-friendly — alex_verem · 2026-07-22
- A better path to agent autonomy is running waves, finding friction, and iterating — JnBrymn · 2026-07-22
- Coding agents are heading toward an AI-writes, AI-reviews, human-approves workflow — aftahi_ai · 2026-07-22
- oMLX 0.5.2 adds Mac menu-bar stats, low-bit decode kernels, and faster downloads — awnihannun · 2026-07-22
- GitHub review bot hits its PR limit and forces a 39-minute cooldown — DanielLockyer · 2026-07-22
- Max reasoning effort appears to be mobile-only in Codex Remote, not desktop — GabGarrett · 2026-07-22