← Search

Andras Geiszl

1 accepted papers

2025

Reward Learning from Multiple Feedback Types

ICLR 2025poster

Learning rewards from preference feedback has become an important tool in the alignment of agentic models. Preference-based feedback, often implemented as a binary comparison between multiple completions, is an established method to acquire large-scale human feedback. However, human feedback in othe…