Tester: Unreleased Astra Solves Every Logic Puzzle Purely by Reasoning, vs 20-30% for sol 5.6

scaling01 · x · 2026-09-06

Blogger @FakePsyho reports testing the rumored (unverified) Astra model with a large batch of logic puzzles — and it solved all of them via pure logical reasoning, no code, including:

For comparison, sol 5.6 managed only 20-30% on similar tests. After analyzing the results, he found every solution legit, and says Astra now explains solving paths better than he can. Whether OpenAI RL'd logic puzzles to death doesn't matter to him — the results are real.

Related event: Rumored Astra Model Outperforms Sol 5.6 on Logic Puzzles(2 posts)→

Original post →

More from Models

Models channel →