← Search

Albert Yu

7 accepted papers

2026

Mixed-Initiative Dialog for Human-Robot Collaborative Manipulation

ICRA 2026poster

Effective robotic systems for long-horizon human-robot collaboration must adapt to a wide range of human partners, whose physical behavior, willingness to assist, and understanding of the robot's capabilities may change over time. This demands a tightly coupled communication loop that grants both ag…

2025

BUMBLE: Unifying Reasoning and Acting with Vision-Language Models for Building-wide Mobile Manipulation

ICRA 2025

To operate at a building scale, service robots must perform long-horizon mobile manipulation tasks by navigating to different rooms, accessing multiple floors, and interacting with a wide and unseen range of everyday objects. We refer to these tasks as Building-wide Mobile Manipulation. To tackle th

Cited by 31SourceScholar
2022

Don’t Start From Scratch: Leveraging Prior Data to Automate Robotic Reinforcement Learning

CoRL 2022poster

Reinforcement learning (RL) algorithms hold the promise of enabling autonomous skill acquisition for robotic systems. However, in practice, real-world robotic RL typically requires time consuming data collection and frequent human intervention to reset the environment. Moreover, robotic policies lea…

Cited by 48SourceScholar
2021

Parrot: Data-Driven Behavioral Priors for Reinforcement Learning

ICLR 2021oral

Reinforcement learning provides a general framework for flexible decision making and control, but requires extensive data collection for each new task that an agent needs to learn. In other machine learning fields, such as natural language processing or computer vision, pre-training on large, previo…

Cited by 168SourcePDFScholar
2020

Chaining Behaviors from Data with Model-Free Reinforcement Learning

CoRL 2020

Reinforcement learning has been applied to a wide variety of robotics problems, but most of such applications involve collecting data from scratch for each new task. Since the amount of robot data we can collect for any single task is limited by time and cost considerations, the learned behavior is

Cited by 0SourcePDFScholar