2023
Multiple-policy High-confidence Policy Evaluation
AISTATS 2023poster
In reinforcement learning applications, we often want to accurately estimate the return of several policies of interest. We study this problem, multiple-policy high-confidence policy evaluation, where the goal is to estimate the return of all given target policies up to a desired accuracy with as fe…