← Search

Daiqi Gao

2 accepted papers

2025

Harnessing Causality in Reinforcement Learning with Bagged Decision Times

AISTATS 2025poster

We consider reinforcement learning (RL) for a class of problems with bagged decision times. A bag contains a finite sequence of consecutive decision times. The transition dynamics are non-Markovian and non-stationary within a bag. All actions within a bag jointly impact a single reward, observed at…

Cited by 0SourceScholar