IJCAI 2022poster3 citations

Multi-policy Grounding and Ensemble Policy Learning for Transfer Learning with Dynamics Mismatch

Hyun-Rok Lee, Ram Ananth Sreenivasan, Yeonjeong Jeong, Jongseong Jang, Dongsub Shim, Chi-Guhn Lee

Abstract

We propose a new transfer learning algorithm between tasks with different dynamics. The proposed algorithm solves an Imitation from Observation problem (IfO) to ground the source environment to the target task before learning an optimal policy in the grounded environment. The learned policy is deployed in the target task without additional training. A particular feature of our algorithm is the employment of multiple rollout policies during training with a goal to ground the environment more globally; hence, it is named as Multi-Policy Grounding (MPG). The quality of final policy is further enhanced via ensemble policy learning. We demonstrate the superiority of the proposed algorithm analytically and numerically. Numerical studies show that the proposed multi-policy approach allows comparable grounding with single policy approach with a fraction of target samples, hence the algorithm is able to maintain the quality of obtained policy even as the number of interactions with the target environment becomes extremely small.

Machine Learning: Multi-task and Transfer LearningMachine Learning: Deep Reinforcement LearningMachine Learning: Ensemble MethodsMachine Learning: Generative Adverserial NetworksMachine Learning: Reinforcement Learning
BibTeX
@inproceedings{ijcai2022p440,
  title     = {Multi-policy Grounding and Ensemble Policy Learning for Transfer Learning with Dynamics Mismatch},
  author    = {Lee, Hyun-Rok and Sreenivasan, Ram Ananth and Jeong, Yeonjeong and Jang, Jongseong and Shim, Dongsub and Lee, Chi-Guhn},
  booktitle = {Proceedings of the Thirty-First International Joint Conference on
               Artificial Intelligence, {IJCAI-22}},
  publisher = {International Joint Conferences on Artificial Intelligence Organization},
  editor    = {Lud De Raedt},
  pages     = {3171--3177},
  year      = {2022},
  month     = {7},
  note      = {Main Track},
  doi       = {10.24963/ijcai.2022/440},
  url       = {https://doi.org/10.24963/ijcai.2022/440},
}
Multi-policy Grounding and Ensemble Policy Learning for Transfer Learning with Dynamics Mismatch · IJCAI 2022