← Search

Ceren Budak

3 accepted papers

2025

Deep Value Benchmark: Measuring Whether Models Generalize Deep values or Shallow Preferences

NeurIPS 2025spotlight

We introduce the Deep Value Benchmark (DVB), an evaluation framework that directly tests whether large language models (LLMs) learn fundamental human values or merely surface-level preferences. This distinction is critical for AI alignment: Systems that capture deeper values are likely to generalize…

Cited by 0SourceScholar
2025

When People are Floods: Analyzing Dehumanizing Metaphors in Immigration Discourse with Large Language Models

ACL 2025long

Metaphor, discussing one concept in terms of another, is abundant in politics and can shape how people understand important issues. We develop a computational approach to measure metaphorical language, focusing on immigration discourse on social media. Grounded in qualitative social science research…

Cited by 0SourcePDFScholar