2020
Ridge Rider: Finding Diverse Solutions by Following Eigenvectors of the Hessian
NeurIPS 2020poster
Over the last decade, a single algorithm has changed many facets of our lives - Stochastic Gradient Descent (SGD). In the era of ever decreasing loss functions, SGD and its various offspring have become the go-to optimization tool in machine learning and are a key component of the success of deep ne…