2022
On global convergence of ResNets: From finite to infinite width using linear parameterization
NeurIPS 2022accept
Overparameterization is a key factor in the absence of convexity to explain global convergence of gradient descent (GD) for neural networks. Beside the well studied lazy regime, infinite width (mean field) analysis has been developed for shallow networks, using on convex optimization technics. To br…