Schema Harness Claims High Scores on ARC-AGI-3

TFenrir · reddit · 2026-07-17

This post introduces a harness called Schema, which, when paired with Fable+4.8 or GPT 5.6 Sol, reportedly achieved scores of 99% and 95.35% on ARC-AGI-3, respectively.

A link to the project homepage is included, clarifying that this is an LLM-centric evaluation/execution framework rather than just a benchmark result for a single model. Key takeaways include:

This serves better as research/evaluation material for those interested in benchmarks and execution frameworks.

Related event: Schema Harness Sparks ARC-AGI-3 Debate(14 posts)→

Original post →

More from coding & agent

coding & agent channel →