← Search

João Sedoc

14 accepted papers

2025

The Illusion of Empathy: How AI Chatbots Shape Conversation Perception

AAAI 2025technical

As AI chatbots increasingly incorporate empathy, understanding user-centered perceptions of chatbot empathy and its impact on conversation quality remains essential yet under-explored. This study examines how chatbot identity and perceived empathy influence users' overall conversation experience. An…

2024

Large Human Language Models: A Need and the Challenges

NAACL 2024long

As research in human-centered NLP advances, there is a growing recognition of the importance of incorporating human and social factors into NLP models. At the same time, our NLP systems have become heavily reliant on LLMs, most of which do not model authors. To build NLP systems that can truly under…

Cited by 9SourcePDFScholar
2024

Modeling Human Subjectivity in LLMs Using Explicit and Implicit Human Factors in Personas

EMNLP 2024finding

Large language models (LLMs) are increasingly being used in human-centered social scientific tasks, such as data annotation, synthetic data creation, and engaging in dialog. However, these tasks are highly subjective and dependent on human factors, such as one’s environment, attitudes, beliefs, and…

Cited by 4SourcePDFScholar
2024

On the Role of Summary Content Units in Text Summarization Evaluation

NAACL 2024short

At the heart of the Pyramid evaluation method for text summarization lie human written summary content units (SCUs). These SCUs areconcise sentences that decompose a summary into small facts. Such SCUs can be used to judge the quality of a candidate summary, possibly partially automated via natural…

2023

A Needle in a Haystack: An Analysis of High-Agreement Workers on MTurk for Summarization

ACL 2023long

To prevent the costly and inefficient use of resources on low-quality annotations, we want a method for creating a pool of dependable annotators who can effectively complete difficult tasks, such as evaluating automatic summarization. Thus, we investigate the recruitment of high-quality Amazon Mecha…

Cited by 11SourcePDFScholar
2023

An Integrative Survey on Mental Health Conversational Agents to Bridge Computer Science and Medical Perspectives

EMNLP 2023long main

Mental health conversational agents (a.k.a. chatbots) are widely studied for their potential to offer accessible support to those experiencing mental health challenges. Previous surveys on the topic primarily consider papers published in either computer science or medicine, leading to a divide in un…

Cited by 0SourcecodeScholar
2023

Common Law Annotations: Investigating the Stability of Dialog System Output Annotations

ACL 2023findings

Metrics for Inter-Annotator Agreement (IAA), like Cohen’s Kappa, are crucial for validating annotated datasets. Although high agreement is often used to show the reliability of annotation procedures, it is insufficient to ensure or reproducibility. While researchers are encouraged to increase annota…

Cited by 5SourcePDFScholar
2023

Linear Connectivity Reveals Generalization Strategies

ICLR 2023poster

In the mode connectivity literature, it is widely accepted that there are common circumstances in which two neural networks, trained similarly on the same data, will maintain loss when interpolated in the weight space. In particular, transfer learning is presumed to ensure the necessary conditions f…

2022

Automatic Document Selection for Efficient Encoder Pretraining

EMNLP 2022main

Building pretrained language models is considered expensive and data-intensive, but must we increase dataset size to achieve better performance? We propose an alternative to larger training sets by automatically identifying smaller yet domain-representative subsets. We extend Cynical Data Selection,…

2022

Inducing Generalizable and Interpretable Lexica

EMNLP 2022finding

Lexica – words and associated scores – are widely used as simple, interpretable, generalizable language features to predict sentiment, emotions, mental health, and personality. They also provide insight into the psychological features behind those moods and traits. Such lexica, historically created…

2022

Measuring the Language of Self-Disclosure across Corpora

ACL 2022findings

Being able to reliably estimate self-disclosure – a key component of friendship and intimacy – from language is important for many psychology studies. We build single-task models on five self-disclosure corpora, but find that these models generalize poorly; the within-domain accuracy of predicted me…

2021

Measuring the ‘I don’t know’ Problem through the Lens of Gricean Quantity

NAACL 2021long

We consider the intrinsic evaluation of neural generative dialog models through the lens of Grice’s Maxims of Conversation (1975). Based on the maxim of Quantity (be informative), we propose Relative Utterance Quantity (RUQ) to diagnose the ‘I don’t know’ problem, in which a dialog system produces g…