2021
Hyperparameter Auto-Tuning in Self-Supervised Robotic Learning
RA-L 2021
Policy optimization in reinforcement learning requires the selection of numerous hyperparameters across different environments. Fixing them incorrectly may negatively impact optimization performance leading notably to insufficient or redundant learning. Insufficient learning (due to convergence to l