2023
Posterior Sampling for Competitive RL: Function Approximation and Partial Observation
NeurIPS 2023poster
This paper investigates posterior sampling algorithms for competitive reinforcement learning (RL) in the context of general function approximations. Focusing on zero-sum Markov games (MGs) under two critical settings, namely self-play and adversarial learning, we first propose the self-play and adve…