← Search

Ming Yu

4 accepted papers

2019

Convergent Policy Optimization for Safe Reinforcement Learning

NeurIPS 2019poster

We study the safe reinforcement learning problem with nonlinear function approximation, where policy optimization is formulated as a constrained optimization problem with both the objective and the constraint being nonconvex functions. For such a problem, we construct a sequence of surrogate convex…