2022
Data-Efficient Pipeline for Offline Reinforcement Learning with Limited Data
NeurIPS 2022accept
Offline reinforcement learning (RL) can be used to improve future performance by leveraging historical data. There exist many different algorithms for offline RL, and it is well recognized that these algorithms, and their hyperparameter settings, can lead to decision policies with substantially diff…