← Search

Aymen Al Marjani

4 accepted papers

2023

On the Complexity of Differentially Private Best-Arm Identification with Fixed Confidence

NeurIPS 2023poster

Best Arm Identification (BAI) problems are progressively used for data-sensitive applications, such as designing adaptive clinical trials, tuning hyper-parameters, and conducting user studies to name a few. Motivated by the data privacy concerns invoked by these applications, we study the problem of…

Cited by 8SourcePDFScholar
2022

Near Instance-Optimal PAC Reinforcement Learning for Deterministic MDPs

NeurIPS 2022accept

In probably approximately correct (PAC) reinforcement learning (RL), an agent is required to identify an $\epsilon$-optimal policy with probability $1-\delta$. While minimax optimal algorithms exist for this problem, its instance-dependent complexity remains elusive in episodic Markov decision proce…

Cited by 23SourcePDFScholar
2021

Adaptive Sampling for Best Policy Identification in Markov Decision Processes

ICML 2021spotlight

We investigate the problem of best-policy identification in discounted Markov Decision Processes (MDPs) when the learner has access to a generative model. The objective is to devise a learning algorithm returning the best policy as early as possible. We first derive a problem-specific lower bound of…

Cited by 34SourcePDFScholar
2021

Navigating to the Best Policy in Markov Decision Processes

NeurIPS 2021poster

We investigate the classical active pure exploration problem in Markov Decision Processes, where the agent sequentially selects actions and, from the resulting system trajectory, aims at identifying the best policy as fast as possible. We propose a problem-dependent lower bound on the average number…

Cited by 36SourcePDFScholar