← Search

Mengdi Wu

1 accepted papers

2022

Finding the Task-Optimal Low-Bit Sub-Distribution in Deep Neural Networks

ICML 2022spotlight

Quantized neural networks typically require smaller memory footprints and lower computation complexity, which is crucial for efficient deployment. However, quantization inevitably leads to a distribution divergence from the original network, which generally degrades the performance. To tackle this i…