2021
Extracting Strong Policies for Robotics Tasks from Zero-Order Trajectory Optimizers
ICLR 2021poster
Solving high-dimensional, continuous robotic tasks is a challenging optimization problem. Model-based methods that rely on zero-order optimizers like the cross-entropy method (CEM) have so far shown strong performance and are considered state-of-the-art in the model-based reinforcement learning comm…