2020
Linear Symmetric Quantization of Neural Networks for Low-precision Integer Hardware
ICLR 2020poster
With the proliferation of specialized neural network processors that operate on low-precision integers, the performance of Deep Neural Network inference becomes increasingly dependent on the result of quantization. Despite plenty of prior work on the quantization of weights or activations for neural…