Astra solve-rate barely improves at max compute, undercutting the 'too smart to throttle' RL theory

zainhas · x · 2026-09-05

A heatmap shared by zainhas shows Astra's solve-rate density barely moves when scaling inference compute from low to max, contradicting teortaxesTex's theory that the model is 'too smart' for unsaturated tasks and spirals off-track at full compute (who also argued RL on the base isn't finished).

If that explanation were true, zainhas argues, you'd expect at least some sparks of brilliance on the right side of the heatmap — there are none: 'this model is so alien.'

Original post →

More from Models

Models channel →