IJCAI 2023poster20 citations

Explainable Reinforcement Learning via a Causal World Model

Zhongwei Yu, Jingqing Ruan, Dengpeng Xing

Abstract

Generating explanations for reinforcement learning (RL) is challenging as actions may produce long-term effects on the future. In this paper, we develop a novel framework for explainable RL by learning a causal world model without prior knowledge of the causal structure of the environment. The model captures the influence of actions, allowing us to interpret the long-term effects of actions through causal chains, which present how actions influence environmental variables and finally lead to rewards. Different from most explanatory models which suffer from low accuracy, our model remains accurate while improving explainability, making it applicable in model-based learning. As a result, we demonstrate that our causal model can serve as the bridge between explainability and learning.

Machine Learning: ML: Explainable/Interpretable machine learningMachine Learning: ML: CausalityMachine Learning: ML: Reinforcement learning
BibTeX
@inproceedings{ijcai2023p505,
  title     = {Explainable Reinforcement Learning via a Causal World Model},
  author    = {Yu, Zhongwei and Ruan, Jingqing and Xing, Dengpeng},
  booktitle = {Proceedings of the Thirty-Second International Joint Conference on
               Artificial Intelligence, {IJCAI-23}},
  publisher = {International Joint Conferences on Artificial Intelligence Organization},
  editor    = {Edith Elkind},
  pages     = {4540--4548},
  year      = {2023},
  month     = {8},
  note      = {Main Track},
  doi       = {10.24963/ijcai.2023/505},
  url       = {https://doi.org/10.24963/ijcai.2023/505},
}
Explainable Reinforcement Learning via a Causal World Model · IJCAI 2023