LLM Sensitivity Challenges in Abusive Language Detection: Instruction-Tuned vs. Human Feedback
The capacity of large language models (LLMs) to understand and distinguish socially unacceptable texts enables them to play a promising role in abusive language detection. However, various factors can affect their sensitivity. In this work, we test whether LLMs have an unintended bias in abusive lan…