← Search

Uddalak Mukherjee

1 accepted papers

2026

Performative Policy Gradient: Optimality in Performative Reinforcement Learning

ICML 2026poster

Post-deployment machine learning algorithms often influence the environments they act in, and thus *shift* the underlying dynamics that the standard reinforcement learning (RL) methods ignore. While designing optimal algorithms in this *performative* setting has recently been studied in supervised l…

Cited by 0SourceScholar