← Search

Lukas Froehlich

1 accepted papers

2022

On-Policy Model Errors in Reinforcement Learning

ICLR 2022poster

Model-free reinforcement learning algorithms can compute policy gradients given sampled environment transitions, but require large amounts of data. In contrast, model-based methods can use the learned model to generate new data, but model errors and bias can render learning unstable or suboptimal. I…

Cited by 10SourcePDFScholar