Optimal Robust Subsidy Policies for Irrational Agent in Principal-Agent MDPs
We investigate a principal-agent problem modeled within a Markov Decision Process, where the principal and the agent have their own rewards. The principal can provide subsidies to influence the agent’s action choices, and the agent’s resulting action policy determines the rewards accrued to the prin…