← Search

Yuying Sun

1 accepted papers

2024

Pragmatic Feature Preferences: Learning Reward-Relevant Preferences from Human Input

ICML 2024poster

Humans use context to specify preferences over behaviors, i.e. their reward functions. Yet, algorithms for inferring reward models from preference data do not take this social learning view into account. Inspired by pragmatic human communication, we study how to extract fine-grained data regarding w…

Cited by 2SourcePDFScholar