PIGDreamer: Privileged Information Guided World Models for Safe Partially Observable Reinforcement Learning
Partial observability presents a significant challenge for safe reinforcement learning, as it impedes the identification of potential risks and rewards. Leveraging specific types of privileged information during training to mitigate the effects of partial observability has yielded notable empirical…