← Search

Tyler Lu

4 accepted papers

2025

Representative Ranking for Deliberation in the Public Sphere

ICML 2025poster

Online comment sections, such as those on news sites or social media, have the potential to foster informal public deliberation, However, this potential is often undermined by the frequency of toxic or low-quality exchanges that occur in these settings. To combat this, platforms increasingly leverag…

Cited by 0SourcePDFScholar
2020

ConQUR: Mitigating Delusional Bias in Deep Q-Learning

ICML 2020poster

Delusional bias is a fundamental source of error in approximate Q-learning. To date, the only techniques that explicitly address delusion require comprehensive search using tabular value estimates. In this paper, we develop efficient methods to mitigate delusional bias by training Q-approximators wi…

2018

Data center cooling using model-predictive control

NeurIPS 2018poster

Despite impressive recent advances in reinforcement learning (RL), its deployment in real-world physical systems is often complicated by unexpected events, limited data, and the potential for expensive failures. In this paper, we describe an application of RL “in the wild” to the task of regulating…

Cited by 256SourcePDFScholar