2019
A Main/Subsidiary Network Framework for Simplifying Binary Neural Networks
CVPR 2019poster
To reduce memory footprint and run-time latency, techniques such as neural net-work pruning and binarization have been explored separately. However, it is un-clear how to combine the best of the two worlds to get extremely small and efficient models. In this paper, we, for the first time, define t…