← Search

Mohammadsadegh Khorasani

2 accepted papers

2025

Efficiently Escaping Saddle Points for Policy Optimization

UAI 2025

Policy gradient (PG) is widely used in reinforcement learning due to its scalability and good performance. In recent years, several variance-reduced PG methods have been proposed with a theoretical guarantee of converging to an approximate first-order stationary point (FOSP) with the sample complexi

2025

Hierarchical Reinforcement Learning with Targeted Causal Interventions

ICML 2025poster

Hierarchical reinforcement learning (HRL) improves the efficiency of long-horizon reinforcement-learning tasks with sparse rewards by decomposing the task into a hierarchy of subgoals. The main challenge of HRL is efficient discovery of the hierarchical structure among subgoals and utilizing this st…

Cited by 0SourcePDFScholar