2021
Reinforcement Learning for Constrained Markov Decision Processes
AISTATS 2021poster
In this paper, we consider the problem of optimization and learning for constrained and multi-objective Markov decision processes, for both discounted rewards and expected average rewards. We formulate the problems as zero-sum games where one player (the agent) solves a Markov decision problem and i…