Grok 4.7 delayed a few more days as Musk blames RL length penalty for giving up early

rohanpaul_ai · x · 2026-09-13

Elon Musk confirmed Grok 4.7 still needs a few more days. He explained the delay: the team may have penalized response length too much in RL, causing the model to give up on hard tasks it can actually solve too early and to lack rigor in checking its own work.

Related event: Musk Says Grok 4.7 Delayed Days by Overly Harsh RL Length Penalty(4 posts)→

Original post →

More from Models

Models channel →