2023
Maximum Entropy Population-Based Training for Zero-Shot Human-AI Coordination
AAAI 2023technical
We study the problem of training a Reinforcement Learning (RL) agent that is collaborative with humans without using human data. Although such agents can be obtained through self-play training, they can suffer significantly from the distributional shift when paired with unencountered partners, such…