2025
Policy Regularization on Globally Accessible States in Cross-Dynamics Reinforcement Learning
ICML 2025spotlight
To learn from data collected in diverse dynamics, Imitation from Observation (IfO) methods leverage expert state trajectories based on the premise that recovering expert state distributions in other dynamics facilitates policy learning in the current one. However, Imitation Learning inherently impos…