2019
Learning to Quantize Deep Networks by Optimizing Quantization Intervals With Task Loss
CVPR 2019oral
Reducing bit-widths of activations and weights of deep networks makes it efficient to compute and store them in memory, which is crucial in their deployments to resource-limited devices, such as mobile phones. However, decreasing bit-widths with quantization generally yields drastically degraded acc…