← Search

Kelly W. Zhang

3 accepted papers

2025

A Deployed Online Reinforcement Learning Algorithm in an Oral Health Clinical Trial

AAAI 2025technical

Dental disease is a prevalent chronic condition associated with substantial financial burden, personal suffering, and increased risk of systemic diseases. Despite widespread recommendations for twice-daily tooth brushing, adherence to recommended oral self-care behaviors remains sub-optimal due to f…

2025

Contextual Thompson Sampling via Generation of Missing Data

NeurIPS 2025poster

We introduce a framework for Thompson sampling (TS) contextual bandit algorithms, in which the algorithm's ability to quantify uncertainty and make decisions depends on the quality of a generative model that is learned offline. Instead of viewing uncertainty in the environment as arising from unobse…

Cited by 0SourceScholar
2023

Reward Design for an Online Reinforcement Learning Algorithm Supporting Oral Self-Care

AAAI 2023technical

While dental disease is largely preventable, professional advice on optimal oral hygiene practices is often forgotten or abandoned by patients. Therefore patients may benefit from timely and personalized encouragement to engage in oral self-care behaviors. In this paper, we develop an online reinfor…