← Search

Skander Karkar

1 accepted papers

2023

Module-wise Training of Neural Networks via the Minimizing Movement Scheme

NeurIPS 2023poster

Greedy layer-wise or module-wise training of neural networks is compelling in constrained and on-device settings where memory is limited, as it circumvents a number of problems of end-to-end back-propagation. However, it suffers from a stagnation problem, whereby early layers overfit and deeper laye…

Cited by 3SourcePDFScholar