2026
Primal-Dual Policy Optimization for Adversarial Linear CMDPs
ICLR 2026poster
Existing work on linear constrained Markov decision processes (CMDPs) has primarily focused on stochastic settings, where the losses and costs are either fixed or drawn from fixed distributions. However, such formulations are inherently vulnerable to adversarially changing environments. To overcome…