← Search

Baher abdulhai

1 accepted papers

2023

Conservative Bayesian Model-Based Value Expansion for Offline Policy Optimization

ICLR 2023poster

Offline reinforcement learning (RL) addresses the problem of learning a performant policy from a fixed batch of data collected by following some behavior policy. Model-based approaches are particularly appealing in the offline setting since they can extract more learning signals from the logged data…