← Search

Haruki Settai

1 accepted papers

2025

A Temporal Difference Method for Stochastic Continuous Dynamics

NeurIPS 2025poster

For continuous systems modeled by dynamical equations such as ODEs and SDEs, Bellman's principle of optimality takes the form of the Hamilton-Jacobi-Bellman (HJB) equation, which provides the theoretical target of reinforcement learning (RL). Although recent advances in RL successfully leverage this…

Cited by 0SourcecodeScholar