← Search

Eric Pacuit

3 accepted papers

2025

Learning to Manipulate Under Limited Information

AAAI 2025technical

By classic results in social choice theory, any reasonable preferential voting method sometimes gives individuals an incentive to report an insincere preference. The extent to which different voting methods are more or less resistant to such strategic manipulation has become a key consideration for…

2024

Position: Social Choice Should Guide AI Alignment in Dealing with Diverse Human Feedback

ICML 2024poster

Foundation models such as GPT-4 are fine-tuned to avoid unsafe or otherwise problematic behavior, such as helping to commit crimes or producing racist text. One approach to fine-tuning, called reinforcement learning from human feedback, learns from humans’ expressed preferences over multiple outputs…

Cited by 29SourcePDFScholar