2024
Absolute Policy Optimization: Enhancing Lower Probability Bound of Performance with High Confidence
ICML 2024poster
In recent years, trust region on-policy reinforcement learning has achieved impressive results in addressing complex control tasks and gaming scenarios. However, contemporary state-of-the-art algorithms within this category primarily emphasize improvement in expected performance, lacking the ability…