2021
One-pass Stochastic Gradient Descent in overparametrized two-layer neural networks
AISTATS 2021poster
There has been a recent surge of interest in understanding the convergence of gradient descent (GD) and stochastic gradient descent (SGD) in overparameterized neural networks. Most previous work assumes that the training data is provided a priori in a batch, while less attention has been paid to the…