2022
A Unifying Theory of Thompson Sampling for Continuous Risk-Averse Bandits
AAAI 2022technical
This paper unifies the design and the analysis of risk-averse Thompson sampling algorithms for the multi-armed bandit problem for a class of risk functionals ρ that are continuous and dominant. We prove generalised concentration bounds for these continuous and dominant risk functionals and show that…