2022
Learning Goal-Conditioned Policies Offline with Self-Supervised Reward Shaping
CoRL 2022poster
Developing agents that can execute multiple skills by learning from pre-collected datasets is an important problem in robotics, where online interaction with the environment is extremely time-consuming. Moreover, manually designing reward functions for every single desired skill is prohibitive. Prio…