2025
Generalized Linear Bandits: Almost Optimal Regret with One-Pass Update
NeurIPS 2025poster
We study the generalized linear bandit (GLB) problem, a contextual multi-armed bandit framework that extends the classical linear model by incorporating a non-linear link function, thereby modeling a broad class of reward distributions such as Bernoulli and Poisson. While GLBs are widely applicable…