← Search

Wesley H. Holliday

4 accepted papers

2025

Learning to Manipulate Under Limited Information

AAAI 2025technical

By classic results in social choice theory, any reasonable preferential voting method sometimes gives individuals an incentive to report an insincere preference. The extent to which different voting methods are more or less resistant to such strategic manipulation has become a key consideration for…

2024

Conditional and Modal Reasoning in Large Language Models

EMNLP 2024main

The reasoning abilities of large language models (LLMs) are the topic of a growing body of research in AI and cognitive science. In this paper, we probe the extent to which twenty-nine LLMs are able to distinguish logically correct inferences from logically fallacious ones. We focus on inference pat…

2024

Position: Social Choice Should Guide AI Alignment in Dealing with Diverse Human Feedback

ICML 2024poster

Foundation models such as GPT-4 are fine-tuned to avoid unsafe or otherwise problematic behavior, such as helping to commit crimes or producing racist text. One approach to fine-tuning, called reinforcement learning from human feedback, learns from humans’ expressed preferences over multiple outputs…

Cited by 29SourcePDFScholar