← Search

Yuming Shao

3 accepted papers

2026

The Pareto-optimal Trade-off between Regret and Statistical Inference in Linear Stochastic Bandits under Safety Constraints

ICML 2026poster

Linear bandits traditionally prioritize regret minimization, often overlooking statistical inference of the underlying parameter as a critical objective. In high-stakes settings such as healthcare, precise parameter estimation is indispensable, as it provides fundamental insights into system mechani…

Cited by 0SourceScholar
2025

Linear Streaming Bandit: Regret Minimization and Fixed-Budget Epsilon-Best Arm Identification

AAAI 2025technical

Recently, there has been a focus on the streaming setting in a line of works on the Multi-Armed Bandit (MAB). In this scenario, a large number of arms arrive in a streaming manner, and the algorithm scans through the stream and stores some arms in its limited processing memory. We advance this line…

Cited by 0SourcePDFScholar