2024
Off-policy Evaluation Beyond Overlap: Sharp Partial Identification Under Smoothness
ICML 2024poster
Off-policy evaluation, and the complementary problem of policy learning, use historical data collected under a logging policy to estimate and/or optimize the value of a target policy. Methods for these tasks typically assume overlap between the target and logging policy, enabling solutions based on…