2024
A Cubic-regularized Policy Newton Algorithm for Reinforcement Learning
AISTATS 2024poster
We consider the problem of control in the setting of reinforcement learning (RL), where model information is not available. Policy gradient algorithms are a popular solution approach for this problem and are usually shown to converge to a stationary point of the value function. In this paper, we pro…