2023
Model-based Reinforcement Learning with Scalable Composite Policy Gradient Estimators
ICML 2023poster
In model-based reinforcement learning (MBRL), policy gradients can be estimated either by derivative-free RL methods, such as likelihood ratio gradients (LR), or by backpropagating through a differentiable model via reparameterization gradients (RP). Instead of using one or the other, the Total Prop…