← Search

Vipul Kumar Sharma

2 accepted papers

2025

Globally Optimal Policy Gradient Algorithms for Reinforcement Learning with PID Control Policies

NeurIPS 2025poster

We develop policy gradient algorithms with global optimality and convergence guarantees for reinforcement learning (RL) with proportional-integral-derivative (PID) parameterized control policies. RL enables learning control policies through direct interaction with a system, without explicit model kn…

Cited by 0SourceScholar
2024

Sampling-based Safe Reinforcement Learning for Nonlinear Dynamical Systems

AISTATS 2024poster

We develop provably safe and convergent reinforcement learning (RL) algorithms for control of nonlinear dynamical systems, bridging the gap between the hard safety guarantees of control theory and the convergence guarantees of RL theory. Recent advances at the intersection of control and RL follow a…