2026
On the Impact of Weight Quantization on Deep Neural Network Uncertainty
AAAI 2026technical
Weight Quantization (WQ) is a key technique for lightweight Deep Neural Network (DNN) computations. While existing algorithms often pursue memory compression and inference acceleration with accuracy comparable to full-precision models, the effect of WQ on DNN uncertainty remains largely unexplored.