← Search

Arman Adibi

4 accepted papers

2024

Stochastic Approximation with Delayed Updates: Finite-Time Rates under Markovian Sampling

AISTATS 2024poster

Motivated by applications in large-scale and multi-agent reinforcement learning, we study the non-asymptotic performance of stochastic approximation (SA) schemes with delayed updates under Markovian sampling. While the effect of delays has been extensively studied for optimization, the manner in whi…

Cited by 14SourcePDFScholar
2022

Collaborative Linear Bandits with Adversarial Agents: Near-Optimal Regret Bounds

NeurIPS 2022accept

We consider a linear stochastic bandit problem involving $M$ agents that can collaborate via a central server to minimize regret. A fraction $\alpha$ of these agents are adversarial and can act arbitrarily, leading to the following tension: while collaboration can potentially reduce regret, it can a…

Cited by 9SourcePDFScholar