2026
Performative Policy Gradient: Optimality in Performative Reinforcement Learning
ICML 2026poster
Post-deployment machine learning algorithms often influence the environments they act in, and thus *shift* the underlying dynamics that the standard reinforcement learning (RL) methods ignore. While designing optimal algorithms in this *performative* setting has recently been studied in supervised l…