IJCAI 2022poster0 citations

Multi-Constraint Deep Reinforcement Learning for Smooth Action Control

Guangyuan Zou, Ying He, F. Richard Yu, Longquan Chen, Weike Pan, Zhong Ming

Abstract

Deep reinforcement learning (DRL) has been studied in a variety of challenging decision-making tasks, e.g., autonomous driving. \textcolor{black}{However, DRL typically suffers from the action shaking problem, which means that agents can select actions with big difference even though states only slightly differ.} One of the crucial reasons for this issue is the inappropriate design of the reward in DRL. In this paper, to address this issue, we propose a novel way to incorporate the smoothness of actions in the reward. Specifically, we introduce sub-rewards and add multiple constraints related to these sub-rewards. In addition, we propose a multi-constraint proximal policy optimization (MCPPO) method to solve the multi-constraint DRL problem. Extensive simulation results show that the proposed MCPPO method has better action smoothness compared with the traditional proportional-integral-differential (PID) and mainstream DRL algorithms. The video is available at https://youtu.be/F2jpaSm7YOg.

Machine Learning: Deep Reinforcement LearningRobotics: Behavior and Control
BibTeX
@inproceedings{ijcai2022p528,
  title     = {Multi-Constraint Deep Reinforcement Learning for Smooth Action Control},
  author    = {Zou, Guangyuan and He, Ying and Yu, F. Richard and Chen, Longquan and Pan, Weike and Ming, Zhong},
  booktitle = {Proceedings of the Thirty-First International Joint Conference on
               Artificial Intelligence, {IJCAI-22}},
  publisher = {International Joint Conferences on Artificial Intelligence Organization},
  editor    = {Lud De Raedt},
  pages     = {3802--3808},
  year      = {2022},
  month     = {7},
  note      = {Main Track},
  doi       = {10.24963/ijcai.2022/528},
  url       = {https://doi.org/10.24963/ijcai.2022/528},
}
Multi-Constraint Deep Reinforcement Learning for Smooth Action Control · IJCAI 2022