2016
Multiple-Play Bandits in the Position-Based Model
NeurIPS 2016poster
Sequentially learning to place items in multi-position displays or lists is a task that can be cast into the multiple-play semi-bandit setting. However, a major concern in this context is when the system cannot decide whether the user feedback for each item is actually exploitable. Indeed, much of t…