← Search

Pierre L’Ecuyer

1 accepted papers

2022

Policy Learning and Evaluation with Randomized Quasi-Monte Carlo

AISTATS 2022poster

Hard integrals arise frequently in reinforcement learning, for example when computing expectations in policy evaluation and policy iteration. They are often analytically intractable and typically estimated with Monte Carlo methods, whose sampling contributes to high variance in policy values and gra…

Cited by 8SourcePDFScholar