2024
Pragmatic Feature Preferences: Learning Reward-Relevant Preferences from Human Input
ICML 2024poster
Humans use context to specify preferences over behaviors, i.e. their reward functions. Yet, algorithms for inferring reward models from preference data do not take this social learning view into account. Inspired by pragmatic human communication, we study how to extract fine-grained data regarding w…