2022
Concurrent Training of a Control Policy and a State Estimator for Dynamic and Robust Legged Locomotion
RA-L 2022
In this letter, we propose a locomotion training framework where a control policy and a state estimator are trained concurrently. The framework consists of a policy network which outputs the desired joint positions and a state estimation network which outputs estimates of the robot’s states such as