← Search

Tim Large

1 accepted papers

2024

Scalable Optimization in the Modular Norm

NeurIPS 2024poster

To improve performance in contemporary deep learning, one is interested in scaling up the neural network in terms of both the number and the size of the layers. When ramping up the width of a single layer, graceful scaling of training has been linked to the need to normalize the weights and their up…