← Search

Junpu Chen

2 accepted papers

2023

Uncertainty-Aware Instance Reweighting for Off-Policy Learning

NeurIPS 2023poster

Off-policy learning, referring to the procedure of policy optimization with access only to logged feedback data, has shown importance in various important real-world applications, such as search engines and recommender systems. While the ground-truth logging policy is usually unknown, previous work…

Cited by 7SourcePDFScholar