← Search

Changlin Han

1 accepted papers

2023

Weighted Policy Constraints for Offline Reinforcement Learning

AAAI 2023technical

Offline reinforcement learning (RL) aims to learn policy from the passively collected offline dataset. Applying existing RL methods on the static dataset straightforwardly will raise distribution shift, causing these unconstrained RL methods to fail. To cope with the distribution shift problem, a co…