2025
MSQ: Memory-Efficient Bit Sparsification Quantization
ICCV 2025poster
As deep neural networks (DNNs) see increased deployment on mobile and edge devices, optimizing model efficiency has become crucial. Mixed-precision quantization is widely favored, as it offers a superior balance between efficiency and accuracy compared to uniform quantization. However, finding the o…