2021
A Joint Imitation-Reinforcement Learning Framework for Reduced Baseline Regret
IROS 2021poster
In various control task domains, existing controllers provide a baseline level of performance that—though possibly suboptimal—should be maintained. Reinforcement learning (RL) algorithms that rely on extensive exploration of the state and action space can be used to optimize a control policy. Howeve…