← Search

David Valensi

1 accepted papers

2024

Tree Search-Based Policy Optimization under Stochastic Execution Delay

ICLR 2024poster

The standard formulation of Markov decision processes (MDPs) assumes that the agent's decisions are executed immediately. However, in numerous realistic applications such as robotics or healthcare, actions are performed with a delay whose value can even be stochastic. In this work, we introduce stoc…