2024
Offline Reward Perturbation Boosts Distributional Shift in Online RL
UAI 2024poster
Offline-to-online reinforcement learning has recently been shown effective in reducing the online sample complexity by first training from offline collected data. However, this additional data source may also invite new poisoning attacks that target offline training. In this work, we reveal such vul…