2024
Trajectory Improvement and Reward Learning from Comparative Language Feedback
CoRL 2024poster
Learning from human feedback has gained traction in fields like robotics and natural language processing in recent years. While prior works mostly rely on human feedback in the form of comparisons, language is a preferable modality that provides more informative insights into user preferences. In th…