← Search

Francesco Vidaich

1 accepted papers

2022

Lifelong Hyper-Policy Optimization with Multiple Importance Sampling Regularization

AAAI 2022technical

Learning in a lifelong setting, where the dynamics continually evolve, is a hard challenge for current reinforcement learning algorithms. Yet this would be a much needed feature for practical applications. In this paper, we propose an approach which learns a hyper-policy, whose input is time, that…