← Search

Pierre-Henri Wuillemin

2 accepted papers

2023

Warm-Starting Nested Rollout Policy Adaptation with Optimal Stopping

AAAI 2023technical

Nested Rollout Policy Adaptation (NRPA) is an approach using online learning policies in a nested structure. It has achieved a great result in a variety of difficult combinatorial optimization problems. In this paper, we propose Meta-NRPA, which combines optimal stopping theory with NRPA for warm-st…

Cited by 9SourcePDFScholar
2021

Learning Continuous High-Dimensional Models using Mutual Information and Copula Bayesian Networks

AAAI 2021technical

We propose a new framework to learn non-parametric graphical models from continuous observational data. Our method is based on concepts from information theory in order to discover independences and causality between variables: the conditional and multivariate mutual information (such as cite{verny2…

Cited by 4SourcePDFScholar