← Search

Jingmin Wang

1 accepted papers

2025

Reflect-then-Plan: Offline Model-Based Planning through a Doubly Bayesian Lens

ICML 2025poster

Offline reinforcement learning (RL) is crucial when online exploration is costly or unsafe but often struggles with high epistemic uncertainty due to limited data. Existing methods rely on fixed conservative policies, restricting adaptivity and generalization. To address this, we propose Reflect-the…

Cited by 0SourcePDFScholar