2020
RaCT: Toward Amortized Ranking-Critical Training For Collaborative Filtering
ICLR 2020poster
We investigate new methods for training collaborative filtering models based on actor-critic reinforcement learning, to more directly maximize ranking-based objective functions. Specifically, we train a critic network to approximate ranking-based metrics, and then update the actor network to directl…