← Search

Yunzong Xu

3 accepted papers

2025

Greedy Algorithms for Structured Bandits: A Sharp Characterization of Asymptotic Success / Failure

NeurIPS 2025poster

We study the greedy (exploitation-only) algorithm in bandit problems with a known reward structure. We allow arbitrary finite reward structures, while prior work focused on a few specific ones. We fully characterize when the greedy algorithm asymptotically succeeds or fails, in the sense of sublinea…

Cited by 0SourceScholar
2020

Online Pricing with Offline Data: Phase Transition and Inverse Square Law

ICML 2020poster

This paper investigates the impact of pre-existing offline data on online learning, in the context of dynamic pricing. We study a single-product dynamic pricing problem over a selling horizon of T periods. The demand in each period is determined by the price of the product according to a linear dema…

Cited by 52SourcePDFScholar