2023
Rethinking Gauss-Newton for learning over-parameterized models
NeurIPS 2023poster
This work studies the global convergence and implicit bias of Gauss Newton's (GN) when optimizing over-parameterized one-hidden layer networks in the mean-field regime. We first establish a global convergence result for GN in the continuous-time limit exhibiting a faster convergence rate compared t…