2023
Delayed Feedback in Generalised Linear Bandits Revisited
AISTATS 2023poster
The stochastic generalised linear bandit is a well-understood model for sequential decision-making problems, with many algorithms achieving near-optimal regret guarantees under immediate feedback. However, the stringent requirement for immediate rewards is unmet in many real-world applications where…