← Search

Sam Lobel*

1 accepted papers

2020

RaCT: Toward Amortized Ranking-Critical Training For Collaborative Filtering

ICLR 2020poster

We investigate new methods for training collaborative filtering models based on actor-critic reinforcement learning, to more directly maximize ranking-based objective functions. Specifically, we train a critic network to approximate ranking-based metrics, and then update the actor network to directl…

Cited by 38SourcecodeScholar