2023
Probing LLMs for hate speech detection: strengths and vulnerabilities
EMNLP 2023long findings
Recently efforts have been made by social media platforms as well as researchers to detect hateful or toxic language using large language models. However, none of these works aim to use explanation, additional context and victim community information in the detection process. We utilise different pr…