2021
Multitask Bandit Learning Through Heterogeneous Feedback Aggregation
AISTATS 2021poster
In many real-world applications, multiple agents seek to learn how to perform highly related yet slightly different tasks in an online bandit learning protocol. We formulate this problem as the $\epsilon$-multi-player multi-armed bandit problem, in which a set of players concurrently interact with a…