2022
Efficient Learning of Multiple NLP Tasks via Collective Weight Factorization on BERT
NAACL 2022findings
The Transformer architecture continues to show remarkable performance gains in many Natural Language Processing tasks. However, obtaining such state-of-the-art performance in different tasks requires fine-tuning the same model separately for each task. Clearly, such an approach is demanding in terms…