2018
Learning a Structured Neural Network Policy for a Hopping Task
RA-L 2018
In this letter, we present a method for learning a reactive policy for a simple dynamic locomotion task involving hard impact and switching contacts where we assume the contact location and contact timing to be unknown. To learn such a policy, we use optimal control to optimize a local controller fo