← Search

Nicolas Chopin

5 accepted papers

2024

Logarithmic Smoothing for Pessimistic Off-Policy Evaluation, Selection and Learning

NeurIPS 2024spotlight

This work investigates the offline formulation of the contextual bandit problem, where the goal is to leverage past interactions collected under a behavior policy to evaluate, select, and learn new, potentially better-performing, policies. Motivated by critical applications, we move beyond point est…

2023

Computational Doob h-transforms for Online Filtering of Discretely Observed Diffusions

ICML 2023poster

This paper is concerned with online filtering of discretely observed nonlinear diffusion processes. Our approach is based on the fully adapted auxiliary particle filter, which involves Doob's $h$-transforms that are typically intractable. We propose a computational framework to approximate these $h$…

Cited by 6SourcePDFScholar