Blog Series Revisits AdaGrad, Reproducing Full Derivation via Upper Bound Minimization

aaron_defazio · x · 2026-09-05

Part IX of the 'Revisiting Convergence Results in Convex Optimization' series takes a fresh look at AdaGrad, the seminal adaptive gradient algorithm. The post reproduces its full derivation via a 'convergence analysis → minimize the upper bound → optimal preconditioning matrix' pipeline, placing adaptive learning-rate methods within a unified convex-optimization convergence framework.

Original post →

More from Research

Research channel →