← Search

Jerry Zhi-Yang He

5 accepted papers

2025

Context Steering: Controllable Personalization at Inference Time

ICLR 2025poster

To deliver high-quality, personalized responses, large language models (LLMs) must effectively incorporate context — personal, demographic, and cultural information specific to an end-user. For example, asking the model to explain Newton's second law with the context "I am a toddler'' should produce…

Cited by 0SourcePDFScholar
2023

Causal Confusion and Reward Misidentification in Preference-Based Reward Learning

ICLR 2023poster

Learning policies via preference-based reward learning is an increasingly popular method for customizing agent behavior, but has been shown anecdotally to be prone to spurious correlations and reward hacking behaviors. While much prior work focuses on causal confusion in reinforcement learning and b…

Cited by 59SourcePDFScholar
2023

Quantifying Assistive Robustness Via the Natural-Adversarial Frontier

CoRL 2023poster

Our ultimate goal is to build robust policies for robots that assist people. What makes this hard is that people can behave unexpectedly at test time, potentially interacting with the robot outside its training distribution and leading to failures. Even just measuring robustness is a challenge. Adve…

Cited by 0SourceScholar
2022

Learning Representations that Enable Generalization in Assistive Tasks

CoRL 2022poster

Recent work in sim2real has successfully enabled robots to act in physical environments by training in simulation with a diverse ``population'' of environments (i.e. domain randomization). In this work, we focus on enabling generalization in \emph{assistive tasks}: tasks in which the robot is acting…

Cited by 35SourceScholar