2017
DSD: Dense-Sparse-Dense Training for Deep Neural Networks
ICLR 2017poster
Modern deep neural networks have a large number of parameters, making them very hard to train. We propose DSD, a dense-sparse-dense training flow, for regularizing deep neural networks and achieving better optimization performance. In the first D (Dense) step, we train a dense network to learn conne…