← Search

Ilai Bistritz

8 accepted papers

2022

Queue Up Your Regrets: Achieving the Dynamic Capacity Region of Multiplayer Bandits

NeurIPS 2022accept

Abstract Consider $N$ cooperative agents such that for $T$ turns, each agent n takes an action $a_{n}$ and receives a stochastic reward $r_{n}\left(a_{1},\ldots,a_{N}\right)$. Agents cannot observe the actions of other agents and do not know even their own reward function. The agents can communicate…

Cited by 5SourcePDFScholar
2021

Online Learning for Load Balancing of Unknown Monotone Resource Allocation Games

ICML 2021spotlight

Consider N players that each uses a mixture of K resources. Each of the players’ reward functions includes a linear pricing term for each resource that is controlled by the game manager. We assume that the game is strongly monotone, so if each player runs gradient descent, the dynamics converge to a…

Cited by 9SourcePDFScholar
2020

My Fair Bandit: Distributed Learning of Max-Min Fairness with Multi-player Bandits

ICML 2020poster

Consider N cooperative but non-communicating players where each plays one out of M arms for T turns. Players have different utilities for each arm, representable as an NxM matrix. These utilities are unknown to the players. In each turn players receive noisy observations of their utility for their s…

Cited by 44SourcePDFScholar