2021
Rethinking Network Pruning – under the Pre-train and Fine-tune Paradigm
NAACL 2021long
Transformer-based pre-trained language models have significantly improved the performance of various natural language processing (NLP) tasks in the recent years. While effective and prevalent, these models are usually prohibitively large for resource-limited deployment scenarios. A thread of researc…