← Search

Haoyu Xiong

8 accepted papers

2026

Blocking the Leakage: Manifold-Aware Gradient Projection for Long-Horizon Test-Time Adaptation

ICML 2026poster

Test-Time Adaptation (TTA) empowers pre-trained models to adapt online to distribution shifts during inference, but such online updates often become unstable in long-horizon deployments. Prevailing approaches attribute this failure to error accumulation from noisy pseudo-labels, relying on heuristic…

Cited by 0SourceScholar
2026

EgoVerse: An Egocentric Human Dataset for Robot Learning from Around the World

RSS 2026poster

Robot learning increasingly depends on large and diverse data, yet robot data collection remains expensive and difficult to scale. Egocentric human data offer a promising alternative by capturing rich manipulation behavior across everyday environments. However, existing human datasets are often limi…

Cited by 0SourceScholar
2025

Vision in Action: Learning Active Perception from Human Demonstrations

CoRL 2025poster

We present Vision in Action (ViA), an active perception system for bimanual robot manipulation. ViA learns task-relevant active perceptual strategies (e.g., searching, tracking, and focusing) directly from human demonstrations. On the hardware side, ViA employs a simple yet effective 6-DoF robotic n…

Cited by 0SourceScholar
2024

Bimanual Dexterity for Complex Tasks

CoRL 2024poster

To train generalist robot policies, machine learning methods often require a substantial amount of expert human teleoperation data. An ideal robot for humans collecting data is one that closely mimics them: bimanual arms and dexterous hands. However, creating such a bimanual teleoperation system wit…

Cited by 21SourcecodeScholar
2024

SPIN: Simultaneous Perception Interaction and Navigation

CVPR 2024poster

While there has been remarkable progress recently in the fields of manipulation and locomotion mobile manipulation remains a long-standing challenge. Compared to locomotion or static manipulation a mobile system must make a diverse range of long-horizon tasks feasible in unstructured and dynamic env…

2024

STAF: Pushing the Boundaries of Test-Time Adaptation towards Practical Noise Scenarios

COLING 2024main

Test-time adaptation (TTA) aims to adapt the neural network to the distribution of the target domain using only unlabeled test data. Most previous TTA methods have achieved success under mild conditions, such as considering only a single or multiple independent static domains. However, in real-world…

2022

RoboTube: Learning Household Manipulation from Human Videos with Simulated Twin Environments

CoRL 2022oral

We aim to build a useful, reproducible, democratized benchmark for learning household robotic manipulation from human videos. To realize this goal, a diverse, high-quality human video dataset curated specifically for robots is desired. To evaluate the learning progress, a simulated twin environment…

Cited by 12SourceScholar
2021

Learning by Watching: Physical Imitation of Manipulation Skills from Human Videos

IROS 2021poster

Learning from visual data opens the potential to accrue a large range of manipulation behaviors by leveraging human demonstrations without specifying each of them mathe-matically, but rather through natural task specification. In this paper, we present Learning by Watching (LbW), an algorithmic fram…

Cited by 91SourceScholar