2026
Convergence of Fast Policy Iteration in Markov Games and Robust MDPs
AAAI 2026technical
Markov games and robust MDPs are closely related models that involve computing a pair of saddle point policies. As part of the long-standing effort to develop efficient algorithms for these models, the Filar-Tolwinski (FT) algorithm has shown considerable promise. As our first contribution, we demon