2022
Bundled Gradients Through Contact Via Randomized Smoothing
RA-L 2022
The empirical success of derivative-free methods in reinforcement learning for planning through contact seems at odds with the perceived fragility of classical gradient-based optimization methods in these domains. What is causing this gap, and how might we use the answer to improve gradient-based me