← Search

Nicolás Benjamín Ocampo

3 accepted papers

2024

Is Safer Better? The Impact of Guardrails on the Argumentative Strength of LLMs in Hate Speech Countering

EMNLP 2024main

The potential effectiveness of counterspeech as a hate speech mitigation strategy is attracting increasing interest in the NLG research community, particularly towards the task of automatically producing it. However, automatically generated responses often lack the argumentative richness which chara…

2023

Playing the Part of the Sharp Bully: Generating Adversarial Examples for Implicit Hate Speech Detection

ACL 2023findings

Research on abusive content detection on social media has primarily focused on explicit forms of hate speech (HS), that are often identifiable by recognizing hateful words and expressions. Messages containing linguistically subtle and implicit forms of hate speech still constitute an open challenge…

Cited by 21SourcePDFScholar
2023

Unmasking the Hidden Meaning: Bridging Implicit and Explicit Hate Speech Embedding Representations

EMNLP 2023short findings

Research on automatic hate speech (HS) detection has mainly focused on identifying explicit forms of hateful expressions on user-generated content. Recently, a few works have started to investigate methods to address more implicit and subtle abusive content. However, despite these efforts, automated…

Cited by 0SourceScholar