2020
Learning Domain Randomization Distributions for Training Robust Locomotion Policies
IROS 2020poster
This paper considers the problem of learning behaviors in simulation without knowledge of the precise dynamical properties of the target robot platform(s). In this context, our learning goal is to mutually maximize task efficacy on each environment considered and generalization across the widest pos…