← Search

Zhipeng Liang

4 accepted papers

2024

Single-Trajectory Distributionally Robust Reinforcement Learning

ICML 2024poster

To mitigate the limitation that the classical reinforcement learning (RL) framework heavily relies on identical training and test environments, Distributionally Robust RL (DRRL) has been proposed to enhance performance across a range of environments, possibly including unknown test environments. As…

Cited by 13SourcePDFScholar
2022

Private Streaming SCO in $\ell_p$ geometry with Applications in High Dimensional Online Decision Making

ICML 2022spotlight

Differentially private (DP) stochastic convex optimization (SCO) is ubiquitous in trustworthy machine learning algorithm design. This paper studies the DP-SCO problem with streaming data sampled from a distribution and arrives sequentially. We also consider the continual release model where paramete…

Cited by 16SourcePDFScholar
2022

UMIX: Improving Importance Weighting for Subpopulation Shift via Uncertainty-Aware Mixup

NeurIPS 2022accept

Subpopulation shift widely exists in many real-world machine learning applications, referring to the training and test distributions containing the same subpopulation groups but varying in subpopulation frequencies. Importance reweighting is a normal way to handle the subpopulation shift issue by im…

2021

Generalized Linear Bandits with Local Differential Privacy

NeurIPS 2021poster

Contextual bandit algorithms are useful in personalized online decision-making. However, many applications such as personalized medicine and online advertising require the utilization of individual-specific information for effective learning, while user's data should remain private from the server d…