2025
Offline Hierarchical Reinforcement Learning via Inverse Optimization
ICLR 2025poster
Hierarchical policies enable strong performance in many sequential decision-making problems, such as those with high-dimensional action spaces, those requiring long-horizon planning, and settings with sparse rewards. However, learning hierarchical policies from static offline datasets presents a si…