← Search

Alkis Sygkounas

1 accepted papers

2025

REvolve: Reward Evolution with Large Language Models using Human Feedback

ICLR 2025poster

Designing effective reward functions is crucial to training reinforcement learning (RL) algorithms. However, this design is non-trivial, even for domain experts, due to the subjective nature of certain tasks that are hard to quantify explicitly. In recent works, large language models (LLMs) have bee…

Cited by 1SourcePDFScholar