LoopSR: Looping Sim-and-Real for Lifelong Policy Adaptation of Legged Robots
Reinforcement Learning (RL) has shown its remarkable and generalizable capability in legged locomotion through sim-to-real transfer. However, while adaptive methods like domain randomization are expected to enhance policy robustness across diverse environments, they potentially compromise the policy