2021
Sharpness-aware Minimization for Efficiently Improving Generalization
ICLR 2021spotlight
In today's heavily overparameterized models, the value of the training loss provides few guarantees on model generalization ability. Indeed, optimizing only the training loss value, as is commonly done, can easily lead to suboptimal model quality. Motivated by the connection between geometry of the…