Agent Harness Aces ARC-AGI-3

naval · x · 2026-07-16

The Impossible Research team introduced an agent harness capable of playing games, writing code, and reasoning "like a physicist."

The post cites [schema] scores: achieving 99% RHAE on the ARC-AGI-3 Public set using Opus 4.8 + Fable 5, and 95.35% using GPT-5.6 Sol. The author's central claim is that this harness organizes LLMs into workflows better suited for high-difficulty reasoning and task execution.

Related event: Schema Harness Sparks ARC-AGI-3 Debate(14 posts)→

Original post →

More from coding & agent

coding & agent channel →