2021
Towards Understanding Learning in Neural Networks with Linear Teachers
ICML 2021spotlight
Can a neural network minimizing cross-entropy learn linearly separable data? Despite progress in the theory of deep learning, this question remains unsolved. Here we prove that SGD globally optimizes this learning problem for a two-layer network with Leaky ReLU activations. The learned network can i…