2022
Data-Free Network Compression via Parametric Non-Uniform Mixed Precision Quantization
CVPR 2022poster
Deep Neural Networks (DNNs) usually have a large number of parameters and consume a huge volume of storage space, which limits the application of DNNs on memory-constrained devices. Network quantization is an appealing way to compress DNNs. However, most of existing quantization methods require the…