← Search

Vineet Goyal

4 accepted papers

2023

Last Switch Dependent Bandits with Monotone Payoff Functions

ICML 2023poster

In a recent work, Laforgue et al. introduce the model of last switch dependent (LSD) bandits, in an attempt to capture nonstationary phenomena induced by the interaction between the player and the environment. Examples include satiation, where consecutive plays of the same action lead to decreased p…

Cited by 4SourcePDFScholar
2017

Beyond Worst-case: A Probabilistic Analysis of Affine Policies in Dynamic Optimization

NeurIPS 2017spotlight

Affine policies (or control) are widely used as a solution approach in dynamic optimization where computing an optimal adjustable solution is usually intractable. While the worst case performance of affine policies can be significantly bad, the empirical performance is observed to be near-optimal fo…

Cited by 17SourcePDFScholar
2016

Assortment Optimization Under the Mallows model

NeurIPS 2016poster

We consider the assortment optimization problem when customer preferences follow a mixture of Mallows distributions. The assortment optimization problem focuses on determining the revenue/profit maximizing subset of products from a large universe of products; it is an important decision that is comm…

Cited by 28SourcePDFScholar