2022
HARALD: Augmenting Hate Speech Data Sets with Real Data
EMNLP 2022finding
The successful completion of the hate speech detection task hinges upon the availability of rich and variable labeled data, which is hard to obtain. In this work, we present a new approach for data augmentation that uses as input real unlabelled data, which is carefully selected from online platform…