← Search

Ashish Harshvardhan

1 accepted papers

2023

Probing LLMs for hate speech detection: strengths and vulnerabilities

EMNLP 2023long findings

Recently efforts have been made by social media platforms as well as researchers to detect hateful or toxic language using large language models. However, none of these works aim to use explanation, additional context and victim community information in the detection process. We utilise different pr…

Cited by 0SourceScholar