2022
Gradient Information Matters in Policy Optimization by Back-propagating through Model
ICLR 2022poster
Model-based reinforcement learning provides an efficient mechanism to find the optimal policy by interacting with the learned environment. In addition to treating the learned environment like a black-box simulator, a more effective way to use the model is to exploit its differentiability. Such metho…