← Search

Javad Azizi

2 accepted papers

2023

Meta-Learning for Simple Regret Minimization

AAAI 2023technical

We develop a meta-learning framework for simple regret minimization in bandits. In this framework, a learning agent interacts with a sequence of bandit tasks, which are sampled i.i.d. from an unknown prior distribution, and learns its meta-parameters to perform better on future tasks. We propose the…

2023

Overcoming Prior Misspecification in Online Learning to Rank

AISTATS 2023poster

The recent literature on online learning to rank (LTR) has established the utility of prior knowledge to Bayesian ranking bandit algorithms. However, a major limitation of existing work is the requirement for the prior used by the algorithm to match the true prior. In this paper, we propose and anal…

Cited by 0SourcePDFScholar