← Search

Alberto Rumi

4 accepted papers

2025

Stochastic Shortest Path with Sparse Adversarial Costs

NeurIPS 2025poster

We study the adversarial Stochastic Shortest Path (SSP) problem with sparse costs under full-information feedback. In the known transition setting, existing bounds based on Online Mirror Descent (OMD) with negative-entropy regularization scale with $\sqrt{\log S A}$, where $SA$ is the size of the st…

Cited by 0SourceScholar
2024

Bandits with Abstention under Expert Advice

NeurIPS 2024poster

We study the classic problem of prediction with expert advice under bandit feedback. Our model assumes that one action, corresponding to the learner's abstention from play, has no reward or loss on every trial. We propose the CBA (Confidence-rated Bandits with Abstentions) algorithm, which exploits…

2024

Best-of-Both-Worlds Algorithms for Linear Contextual Bandits

AISTATS 2024poster

We study best-of-both-worlds algorithms for $K$-armed linear contextual bandits. Our algorithms deliver near-optimal regret bounds in both the adversarial and stochastic regimes, without prior knowledge about the environment. In the stochastic regime, we achieve the polylogarithmic rate $\frac{(dK)^…

Cited by 6SourcePDFScholar