2026
Efficient Best-of-Both-Worlds Algorithms for Contextual Combinatorial Semi-Bandits
ICLR 2026poster
We introduce the first best-of-both-worlds algorithm for contextual combinatorial semi-bandits that simultaneously guarantees $\widetilde{\mathcal{O}}(\sqrt{T})$ regret in the adversarial regime and $\widetilde{\mathcal{O}}(\ln T)$ regret in the corrupted stochastic regime. Our approach builds on th…