2026
A Reward-Free Viewpoint on Multi-Objective Reinforcement Learning
ICLR 2026poster
Many sequential decision-making tasks involve optimizing multiple conflicting objectives, requiring policies that adapt to different user preferences. In multi-objective reinforcement learning (MORL), one widely studied approach addresses this by training a single policy network conditioned on prefe…