2020
Go Simple and Pre-Train on Domain-Specific Corpora: On the Role of Training Data for Text Classification
COLING 2020main
Pre-trained language models provide the foundations for state-of-the-art performance across a wide range of natural language processing tasks, including text classification. However, most classification datasets assume a large amount labeled data, which is commonly not the case in practical settings…