2023
Meta-Learning Adversarial Bandit Algorithms
NeurIPS 2023poster
We study online meta-learning with bandit feedback, with the goal of improving performance across multiple tasks if they are similar according to some natural similarity measure. As the first to target the adversarial online-within-online partial-information setting, we design meta-algorithms that…