2025
REvolve: Reward Evolution with Large Language Models using Human Feedback
ICLR 2025poster
Designing effective reward functions is crucial to training reinforcement learning (RL) algorithms. However, this design is non-trivial, even for domain experts, due to the subjective nature of certain tasks that are hard to quantify explicitly. In recent works, large language models (LLMs) have bee…