← Search

Xing Qian

1 accepted papers

2024

HateModerate: Testing Hate Speech Detectors against Content Moderation Policies

NAACL 2024findings

To protect users from massive hateful content, existing works studied automated hate speech detection. Despite the existing efforts, one question remains: Do automated hate speech detectors conform to social media content policies? A platform’s content policies are a checklist of content moderated b…