2024
Privacy-Constrained Policies via Mutual Information Regularized Policy Gradients
AISTATS 2024poster
As reinforcement learning techniques are increasingly applied to real-world decision problems, attention has turned to how these algorithms use potentially sensitive information. We consider the task of training a policy that maximizes reward while minimizing disclosure of certain sensitive state va…