30 years on, deep learning still lacks a clear answer on flat minima

fleetwood___ · x · 2026-07-29

A post marking 30 years since Schmidhuber’s flat-minima paper says the field still doesn’t know whether flat minima generalize better.

The attached figure contrasts a “flat” and a “sharp” minimum, highlighting a long-running optimization question in deep learning: whether the geometry of the loss basin actually predicts generalization performance.

Original post →

More from Research

Research channel →