2026
Reward Evolution with Graph-Of-Thoughts: A Bi-Level Language Model Framework for Reinforcement Learning
ICRA 2026poster
Designing effective reward functions remains a major challenge in reinforcement learning (RL), often requiring considerable human expertise and iterative refinement. Recent advances leverage Large Language Models (LLMs) for automated reward design, but these approaches are limited by hallucinations,…