2024
Policy Rehearsing: Training Generalizable Policies for Reinforcement Learning
ICLR 2024poster
Human beings can make adaptive decisions in a preparatory manner, i.e., by making preparations in advance, which offers significant advantages in scenarios where both online and offline experiences are expensive and limited. Meanwhile, current reinforcement learning methods commonly rely on numerous…