Astra appears to solve competition math without verbalized reasoning, monitorability worsens

sjgadler · x · 2026-09-04

Ryan Greenblatt argues GPT-6 Astra shows a massive jump in opaque reasoning: benchmarks suggest it can solve hard competition math entirely in its head, where prior AIs handled only basic word problems — with caveats around possible contamination. UK AISI separately found Astra has much worse monitorability. He suspects architectural changes with increased serial depth, though a normal pretraining scale-up is plausible. Nathan Calvin adds the monitorability situation is worse than expected and harder to coordinate on than a single taboo.

Related event: Researchers Warn GPT-6 Astra Is Far Less Monitorable Despite Capability Jump(11 posts)→

Original post →

More from Models

Models channel →