2020
A Bandit Learning Algorithm and Applications to Auction Design
NeurIPS 2020spotlight
We consider online bandit learning in which at every time step, an algorithm has to make a decision and then observe only its reward. The goal is to design efficient (polynomial-time) algorithms that achieve a total reward approximately close to that of the best fixed decision in hindsight. In thi…