← Search

Jingyuan Deng

3 accepted papers

2021

Incentivized Bandit Learning with Self-Reinforcing User Preferences

ICML 2021spotlight

In this paper, we investigate a new multi-armed bandit (MAB) online learning model that considers real-world phenomena in many recommender systems: (i) the learning agent cannot pull the arms by itself and thus has to offer rewards to users to incentivize arm-pulling indirectly; and (ii) if users wi…

Cited by 6SourcePDFScholar