2022
Koopman Q-learning: Offline Reinforcement Learning via Symmetries of Dynamics
ICML 2022spotlight
Offline reinforcement learning leverages large datasets to train policies without interactions with the environment. The learned policies may then be deployed in real-world settings where interactions are costly or dangerous. Current algorithms over-fit to the training dataset and as a consequence p…