2020
On Thompson Sampling for Smoother-than-Lipschitz Bandits
AISTATS 2020poster
Thompson Sampling is a well established approach to bandit and reinforcement learning problems. However its use in continuum armed bandit problems has received relatively little attention. We provide the first bounds on the regret of Thompson Sampling for continuum armed bandits under weak conditio…