← Search

Tianrui Chen

2 accepted papers

2022

Strategies for Safe Multi-Armed Bandits with Logarithmic Regret and Risk

ICML 2022spotlight

We investigate a natural but surprisingly unstudied approach to the multi-armed bandit problem under safety risk constraints. Each arm is associated with an unknown law on safety risks and rewards, and the learner’s goal is to maximise reward whilst not playing unsafe arms, as determined by a given…

Cited by 14SourcePDFScholar