2025
Online Reinforcement Learning in Non-Stationary Context-Driven Environments
ICLR 2025spotlight
We study online reinforcement learning (RL) in non-stationary environments, where a time-varying exogenous context process affects the environment dynamics. Online RL is challenging in such environments due to "catastrophic forgetting" (CF). The agent tends to forget prior knowledge as it trains on…