2022
Policy Learning and Evaluation with Randomized Quasi-Monte Carlo
AISTATS 2022poster
Hard integrals arise frequently in reinforcement learning, for example when computing expectations in policy evaluation and policy iteration. They are often analytically intractable and typically estimated with Monte Carlo methods, whose sampling contributes to high variance in policy values and gra…