2025
Reflect-then-Plan: Offline Model-Based Planning through a Doubly Bayesian Lens
ICML 2025poster
Offline reinforcement learning (RL) is crucial when online exploration is costly or unsafe but often struggles with high epistemic uncertainty due to limited data. Existing methods rely on fixed conservative policies, restricting adaptivity and generalization. To address this, we propose Reflect-the…