SoftServe preprint brings scalable quasi-Newton optimization to deep learning, beating Adam, Muon and SOAP on ill-conditioned tasks

dianarycai · x · 2026-10-02

Researchers including Diana Cai released SoftServe, a family of quasi-Newton optimization methods designed for non-convex deep learning objectives that scales to very large networks.

Preprint: arXiv:2610.02182.

Related event: SoftServe: Quasi-Newton Optimization Scales to Massive Neural Networks(3 posts)→

Original post →

More from Research

Research channel →