← Search

Maciej Szymczak

1 accepted papers

2020

The Break-Even Point on Optimization Trajectories of Deep Neural Networks

ICLR 2020spotlight

The early phase of training of deep neural networks is critical for their final performance. In this work, we study how the hyperparameters of stochastic gradient descent (SGD) used in the early phase of training affect the rest of the optimization trajectory. We argue for the existence of the "``br…

Cited by 191SourceScholar