← Search

Wenhai Cui

3 accepted papers

2024

Responsible Bandit Learning via Privacy-Protected Mean-Volatility Utility

AAAI 2024technical

For ensuring the safety of users by protecting the privacy, the traditional privacy-preserving bandit algorithm aiming to maximize the mean reward has been widely studied in scenarios such as online ride-hailing, advertising recommendations, and personalized healthcare. However, classical bandit le…

Cited by 1SourcePDFScholar
2023

Opposite Online Learning via Sequentially Integrated Stochastic Gradient Descent Estimators

AAAI 2023technical

Stochastic gradient descent algorithm (SGD) has been popular in various fields of artificial intelligence as well as a prototype of online learning algorithms. This article proposes a novel and general framework of one-sided testing for streaming data based on SGD, which determines whether the u…

Cited by 2SourcePDFScholar