Astra's Post-Training Is Far From Done: More Compute Doesn't Monotonically Help Yet

teortaxesTex · x · 2026-09-04

teortaxesTex notes several Astra benchmark charts reflect the same phenomenon: its post-training is far from finished — the model doesn't yet reliably or monotonically convert more compute into better results, suggesting this skill hasn't been trained in yet.

Related event: Report: GPT-6 Astra post-training unfinished, compute-to-performance scaling still unstable(2 posts)→

Original post →

More from Models

Models channel →