Stochastic Semi-Gradient Descent for Learning Mean Field Games with Population-Aware Function Approximation
Mean field games (MFGs) model interactions in large-population multi-agent systems through population distributions. Traditional learning methods for MFGs are based on fixed-point iteration (FPI), where policy updates and induced population distributions are computed separately and sequentially. How…