2026
Strict Subgoal Execution: Reliable Long-Horizon Planning in Hierarchical Reinforcement Learning
ICLR 2026poster
Long-horizon goal-conditioned tasks pose fundamental challenges for reinforcement learning (RL), particularly when goals are distant and rewards are sparse. While hierarchical and graph-based methods offer partial solutions, their reliance on conventional hindsight relabeling often fails to correct…