← Search

Chongchong Li

1 accepted papers

2022

Gradient Information Matters in Policy Optimization by Back-propagating through Model

ICLR 2022poster

Model-based reinforcement learning provides an efficient mechanism to find the optimal policy by interacting with the learned environment. In addition to treating the learned environment like a black-box simulator, a more effective way to use the model is to exploit its differentiability. Such metho…