← Search

Matthew S. Zhang

1 accepted papers

2022

Convergence and Optimality of Policy Gradient Methods in Weakly Smooth Settings

AAAI 2022technical

Policy gradient methods have been frequently applied to problems in control and reinforcement learning with great success, yet existing convergence analysis still relies on non-intuitive, impractical and often opaque conditions. In particular, existing rates are achieved in limited settings, under s…

Cited by 9SourcePDFScholar