Muon massively beats Adam on tiny superhuman agents across games, drones, driving

jsuarez · x · 2026-08-29

Researcher jsuarez clarifies their work is not on LLMs: they train tiny superhuman agents across games, drones, driving, and commerce sims. Under a full hyperparameter sweep for both optimizers, Muon is a massive improvement over Adam on these tasks. The thread context is a hypothesis about mode-covering in cross-entropy pretraining vs RL exploration.

Related event: Muon Optimizer Beats Adam on Small Superhuman Agents(2 posts)→

Original post →

More from Research

Research channel →