2023
BRExIt: On Opponent Modelling in Expert Iteration
IJCAI 2023poster
Finding a best response policy is a central objective in game theory and multi-agent learning, with modern population-based training approaches employing reinforcement learning algorithms as best-response oracles to improve play against candidate opponents (typically previously learnt policies). We…