2016
Online Stochastic Linear Optimization under One-bit Feedback
ICML 2016poster
In this paper, we study a special bandit setting of online stochastic linear optimization, where only one-bit of information is revealed to the learner at each round. This problem has found many applications including online advertisement and online recommendation. We assume the binary feedback is a…