2025
BAMDP Shaping: a Unified Framework for Intrinsic Motivation and Reward Shaping
ICLR 2025poster
Intrinsic motivation and reward shaping guide reinforcement learning (RL) agents by adding pseudo-rewards, which can lead to useful emergent behaviors. However, they can also encourage counterproductive exploits, e.g., fixation with noisy TV screens. Here we provide a theoretical model which anticip…