← Search

Jashan Shewakramani

1 accepted papers

2021

Permute, Quantize, and Fine-Tune: Efficient Compression of Neural Networks

CVPR 2021poster

Compressing large neural networks is an important step for their deployment in resource-constrained computational platforms. In this context, vector quantization is an appealing framework that expresses multiple parameters using a single code, and has recently achieved state-of-the-art network compr…

Cited by 49PDFcodeScholar