Grok 4.7 delayed a few more days as Musk blames RL length penalty for giving up early
rohanpaul_ai · x · 2026-09-13
Elon Musk confirmed Grok 4.7 still needs a few more days. He explained the delay: the team may have penalized response length too much in RL, causing the model to give up on hard tasks it can actually solve too early and to lack rigor in checking its own work.
Related event: Musk Says Grok 4.7 Delayed Days by Overly Harsh RL Length Penalty(4 posts)→
More from Models
- Mystery 'Kimi Pluto v1' model card spotted on Fireworks, hinting at new Moonshot model — Severe_Post_2751 · 2026-09-13
- User mulls leaving Gemini, asks for honest Claude vs. ChatGPT comparison for B2B work — spacedoutcowboy1 · 2026-09-13
- Reddit post lists why ChatGPT falls short of AGI: vision hallucinations, three hands, refused dilemmas — kaljakin · 2026-09-13
- Rumors Swirl That Gemini 4 Finished Pretraining Early After GDM Discoveries — teortaxesTex · 2026-09-13
- Qwen on M2 Ultra: latest oMLX update brings substantial local inference speedup — Thrumpwart · 2026-09-13
- Timelines Flooded With GPT-6 Astra Robot Demos as Physical AI Hypes Up — CyberRobooo · 2026-09-13