2020
Fast Adaptation to New Environments via Policy-Dynamics Value Functions
ICML 2020poster
Standard RL algorithms assume fixed environment dynamics and require a significant amount of interaction to adapt to new environments. We introduce Policy-Dynamics Value Functions (PD-VF), a novel approach for rapidly adapting to dynamics different from those previously seen in training. PD-VF expli…