2026
Instance-Dependent Continuous-Time Reinforcement Learning via Maximum Likelihood Estimation
ICML 2026poster
Continuous-time reinforcement learning (CTRL) provides a natural framework for sequential decision-making in dynamic environments where interactions evolve continuously over time. While CTRL has shown growing empirical success, its ability to adapt to varying levels of problem difficulty remains poo…