Is Safer Better? The Impact of Guardrails on the Argumentative Strength of LLMs in Hate Speech Countering
The potential effectiveness of counterspeech as a hate speech mitigation strategy is attracting increasing interest in the NLG research community, particularly towards the task of automatically producing it. However, automatically generated responses often lack the argumentative richness which chara…