← Search

Azin Ashkan

4 accepted papers

2015

Cascading Bandits: Learning to Rank in the Cascade Model

ICML 2015poster

A search engine usually outputs a list of K web pages. The user examines this list, from the first web page to the last, and chooses the first attractive page. This model of user behavior is known as the cascade model. In this paper, we propose cascading bandits, a learning variant of the cascade mo…

Cited by 338SourcePDFScholar
2015

Tight Regret Bounds for Stochastic Combinatorial Semi-Bandits

AISTATS 2015poster

A stochastic combinatorial semi-bandit is an online learning problem where at each step a learning agent chooses a subset of ground items subject to constraints, and then observes stochastic weights of these items and receives their sum as a payoff. In this paper, we close the problem of computation…

Cited by 361SourcePDFScholar