← Search

Helen Margetts

2 accepted papers

2021

HateCheck: Functional Tests for Hate Speech Detection Models

ACL 2021long

Detecting online hate is a difficult task that even state-of-the-art models struggle with. Typically, hate speech detection models are evaluated by measuring their performance on held-out test data using metrics such as accuracy and F1 score. However, this approach makes it difficult to identify spe…

2021

Introducing CAD: the Contextual Abuse Dataset

NAACL 2021long

Online abuse can inflict harm on users and communities, making online spaces unsafe and toxic. Progress in automatically detecting and classifying abusive content is often held back by the lack of high quality and detailed datasets. We introduce a new dataset of primarily English Reddit entries whic…