← Search

Jeremy Tien

2 accepted papers

2024

Trajectory Improvement and Reward Learning from Comparative Language Feedback

CoRL 2024poster

Learning from human feedback has gained traction in fields like robotics and natural language processing in recent years. While prior works mostly rely on human feedback in the form of comparisons, language is a preferable modality that provides more informative insights into user preferences. In th…

Cited by 8SourceScholar
2023

Causal Confusion and Reward Misidentification in Preference-Based Reward Learning

ICLR 2023poster

Learning policies via preference-based reward learning is an increasingly popular method for customizing agent behavior, but has been shown anecdotally to be prone to spurious correlations and reward hacking behaviors. While much prior work focuses on causal confusion in reinforcement learning and b…

Cited by 59SourcePDFScholar