Schema Harness Claims High Score on ARC-3
we_are_mammals · reddit · 2026-07-17
The post introduces a new harness named Schema, which claims to achieve exceptionally high scores on the ARC-AGI-3 Public set without modifying model weights. It does so by altering observation modeling, history verification, and plan execution/revision processes.
Reported results include hitting 99% with Claude Opus 4.8 and Fable 5, and 95.35% with GPT-5.6 Sol. The author details a fixed fallback rule: run the stronger configuration first, and if a level falls below a threshold, rerun it with an even stronger config, keeping the highest score per level.
Related event: Schema Harness Sparks ARC-AGI-3 Debate(14 posts)→
More from coding & agent
- Chaining dependent MCP tool calls: no rollback, duplicate risk — agentrsdg · 2026-09-11
- DeepMind-led paper makes design docs the source of truth, code disposable — SMART regenerates in 1.5-3h for ~$100 — Roger_M_Taylor · 2026-09-11
- Agent-built classifier labels 192k docs for $0.70 vs $13-26 with frontier LLMs — vanstriendaniel · 2026-09-11
- MathModelAgent gains traction: auto-solves math modeling and writes a submission-ready paper — jihe520 · 2026-09-11
- alphaXiv open-sources OpenResearch to run parallel research agents with any model — alphaXiv · 2026-09-11
- DeskcommCRM: open-source AI sales CRM with native agents and WhatsApp hits 1k stars — melgarafael · 2026-09-11