← Search

Julia Mendelsohn

6 accepted papers

2025

AI-LieDar : Examine the Trade-off Between Utility and Truthfulness in LLM Agents

NAACL 2025long

Truthfulness (adherence to factual accuracy) and utility (satisfying human needs and instructions) are both fundamental aspects of Large Language Models, yet these goals often conflict (e.g., sell a car with known flaws), making it challenging to achieve both in real-world deployments. We propose AI…

2025

When People are Floods: Analyzing Dehumanizing Metaphors in Immigration Discourse with Large Language Models

ACL 2025long

Metaphor, discussing one concept in terms of another, is abundant in politics and can shape how people understand important issues. We develop a computational approach to measure metaphorical language, focusing on immigration discourse on social media. Grounded in qualitative social science research…

Cited by 0SourcePDFScholar
2023

From Dogwhistles to Bullhorns: Unveiling Coded Rhetoric with Language Models

ACL 2023long

Dogwhistles are coded expressions that simultaneously convey one meaning to a broad audience and a second, often hateful or provocative, meaning to a narrow in-group; they are deployed to evade both political repercussions and algorithmic content moderation. For example, the word “cosmopolitan” in a…

Cited by 23SourcePDFScholar
2022

Challenges and Opportunities in Information Manipulation Detection: An Examination of Wartime Russian Media

EMNLP 2022finding

NLP research on public opinion manipulation campaigns has primarily focused on detecting overt strategies such as fake news and disinformation. However, information manipulation in the ongoing Russia-Ukraine war exemplifies how governments and media also employ more nuanced strategies. We release a…

2021

Detecting Community Sensitive Norm Violations in Online Conversations

EMNLP 2021finding

Online platforms and communities establish their own norms that govern what behavior is acceptable within the community. Substantial effort in NLP has focused on identifying unacceptable behaviors and, recently, on forecasting them before they occur. However, these efforts have largely focused on to…