2018
Weightless: Lossy Weight Encoding For Deep Neural Network Compression
ICLR 2018workshop
The large memory requirements of deep neural networks strain the capabilities of many devices, limiting their deployment and adoption. Model compression methods effectively reduce the memory requirements of these models, usually through applying transformations such as weight pruning or quantization…