2019
Differentiable Soft Quantization: Bridging Full-Precision and Low-Bit Neural Networks
ICCV 2019poster
Hardware-friendly network quantization (e.g., binary/uniform quantization) can efficiently accelerate the inference and meanwhile reduce memory consumption of the deep neural networks, which is crucial for model deployment on resource-limited devices like mobile phones. However, due to the discreten…