← Search

Manuel Tonneau

4 accepted papers

2025

HateDay: Insights from a Global Hate Speech Dataset Representative of a Day on Twitter

ACL 2025long

To address the global challenge of online hate speech, prior research has developed detection models to flag such content on social media. However, due to systematic biases in evaluation datasets, the real-world effectiveness of these models remains unclear, particularly across geographies. We intro…

Cited by 0SourcePDFScholar
2025

When Claims Evolve: Evaluating and Enhancing the Robustness of Embedding Models Against Misinformation Edits

ACL 2025finding

Online misinformation remains a critical challenge, and fact-checkers increasingly rely on claim matching systems that use sentence embedding models to retrieve relevant fact-checks. However, as users interact with claims online, they often introduce edits, and it remains unclear whether current emb…

2024

NaijaHate: Evaluating Hate Speech Detection on Nigerian Twitter Using Representative Data

ACL 2024long

To address the global issue of online hate, hate speech detection (HSD) systems are typically developed on datasets from the United States, thereby failing to generalize to English dialects from the Majority World. Furthermore, HSD models are often evaluated on non-representative samples, raising co…

2022

Multilingual Detection of Personal Employment Status on Twitter

ACL 2022long

Detecting disclosures of individuals’ employment status on social media can provide valuable information to match job seekers with suitable vacancies, offer social protection, or measure labor market flows. However, identifying such personal disclosures is a challenging task due to their rarity in a…