← Search

Dennis Maximilian Elbrächter

1 accepted papers

2019

How degenerate is the parametrization of neural networks with the ReLU activation function?

NeurIPS 2019poster

Neural network training is usually accomplished by solving a non-convex optimization problem using stochastic gradient descent. Although one optimizes over the networks parameters, the main loss function generally only depends on the realization of the neural network, i.e. the function it computes.…

Cited by 43SourcePDFScholar