← Search

Yuchen Xiao

11 accepted papers

2026

Offline Multi-agent Continual Cooperation via Skill Partition and Reuse

ICML 2026poster

Extracting skills from multi-agent offline dataset improves learning efficiency via sharing task-invariant coordination skills among tasks. In settings where tasks occur sequentially and the space of skills grows exponentially, existing approaches that rely on heuristically designed and fixed-sized …

Cited by 0SourceScholar
2026

OmniXtreme: Breaking the Generality Barrier in High-Dynamic Humanoid Control

RSS 2026poster

High-fidelity motion tracking serves as the ultimate litmus test for generalizable, human-level motor skills. However, current policies often hit a “generality barrier”: as motion libraries scale in diversity, tracking fidelity inevitably collapses—especially for real-world deployment of high-dynami…

Cited by 0SourceScholar
2022

A Deeper Understanding of State-Based Critics in Multi-Agent Reinforcement Learning

AAAI 2022technical

Centralized Training for Decentralized Execution, where training is done in a centralized offline fashion, has become a popular solution paradigm in Multi-Agent Reinforcement Learning. Many such methods take the form of actor-critic with state-based critics, since centralized training allows access…

2020

Learning Multi-Robot Decentralized Macro-Action-Based Policies via a Centralized Q-Net

ICRA 2020poster

In many real-world multi-robot tasks, high-quality solutions often require a team of robots to perform asynchronous actions under decentralized control. Decentralized multi-agent reinforcement learning methods have difficulty learning decentralized policies because of the environment appearing to be…

Cited by 39SourceScholar
2019

Online Planning for Target Object Search in Clutter under Partial Observability

ICRA 2019poster

The problem of finding and grasping a target object in a cluttered, uncertain environment, target object search, is a common and important problem in robotics. One key challenge is the uncertainty of locating and recognizing each object in a cluttered environment due to noisy perception and occlusio…

Cited by 89SourceScholar
2018

Near-Optimal Adversarial Policy Switching for Decentralized Asynchronous Multi-Agent Systems

ICRA 2018poster

A key challenge in multi-robot and multi-agent systems is generating solutions that are robust to other self-interested or even adversarial parties who actively try to prevent the agents from achieving their goals. The practicality of existing works addressing this challenge is limited to only small…

Cited by 16SourceScholar
2016

Contact localization through spatially overlapping piezoresistive signals

IROS 2016poster

Achieving high spatial resolution in contact sensing for robotic manipulation often comes at the price of increased complexity in fabrication and integration. One traditional approach is to fabricate a large number of taxels, each delivering an individual, isolated response to a stimulus. In contras…

Cited by 12SourceScholar
2016

On the feasibility of wearable exotendon networks for whole-hand movement patterns in stroke patients

ICRA 2016

Fully wearable hand rehabilitation and assistive devices could extend training and improve quality of life for patients affected by hand impairments. However, such devices must deliver meaningful manipulation capabilities in a small and lightweight package. In this context, this paper investigates t

Cited by 25SourceScholar