2024
Policy Optimization by Looking Ahead for Model-based Offline Reinforcement Learning
ICRA 2024poster
Offline reinforcement learning (RL) aims to optimize a policy, based on pre-collected data, to maximize the cumulative rewards after performing a sequence of actions. Existing approaches learn a value function from historical data and then guide the updating of the policy parameters by maximizing th…