← Search

Akash Mondal

1 accepted papers

2024

A Cubic-regularized Policy Newton Algorithm for Reinforcement Learning

AISTATS 2024poster

We consider the problem of control in the setting of reinforcement learning (RL), where model information is not available. Policy gradient algorithms are a popular solution approach for this problem and are usually shown to converge to a stationary point of the value function. In this paper, we pro…

Cited by 3SourcePDFScholar