2023
TS-UCB: Improving on Thompson Sampling With Little to No Additional Computation
AISTATS 2023poster
Thompson sampling has become a ubiquitous approach to online decision problems with bandit feedback. The key algorithmic task for Thompson sampling is drawing a sample from the posterior of the optimal action. We propose an alternative arm selection rule we dub TS-UCB, that requires negligible addit…