← Search

Yuxuan Wu

8 accepted papers

2026

TrajBooster: Boosting Humanoid Whole-Body Manipulation Via Trajectory-Centric Learning

ICRA 2026poster

Recent Vision-Language-Action (VLA) models show potential to generalize across embodiments but struggle to quickly align with a new robot’s action space when high-quality demonstrations are scarce, especially for bipedal humanoids. We present TrajBooster, a cross-embodiment framework that leverages …

2026

When would Vision-Proprioception Policies Fail in Robotic Manipulation?

ICLR 2026poster

Proprioceptive information is critical for precise servo control by providing real-time robotic states. Its collaboration with vision is highly expected to enhance performances of the manipulation policy in complex tasks. However, recent studies have reported inconsistent observations on the general…

Cited by 0SourcecodeScholar
2025

Bi-Directional Multi-Scale Graph Dataset Condensation via Information Bottleneck

AAAI 2025technical

Dataset condensation has significantly improved model training efficiency, but its application on devices with different computing power brings new requirements for different data sizes. For sparse graph data with non-Euclidean structures, repeated condensation of each scale may lead to significant…

2025

RL-GSBridge: 3D Gaussian Splatting Based Real2Sim2Real Method for Robotic Manipulation Learning

ICRA 2025

Sim-to-Real refers to the process of transferring policies learned in simulation to the real world, which is crucial for achieving practical robotics applications. However, recent Sim2real methods either rely on a large amount of augmented data or large learning models, which is inefficient for spec

Cited by 2SourceScholar
2025

Unsupervised Disentanglement of Content and Style via Variance-Invariance Constraints

ICLR 2025poster

We contribute an unsupervised method that effectively learns disentangled content and style representations from sequences of observations. Unlike most disentanglement algorithms that rely on domain-specific labels or knowledge, our method is based on the insight of domain-general statistical differ…

Cited by 0SourcePDFScholar
2023

Multi-Stream Representation Learning for Pedestrian Trajectory Prediction

AAAI 2023technical

Forecasting the future trajectory of pedestrians is an important task in computer vision with a range of applications, from security cameras to autonomous driving. It is very challenging because pedestrians not only move individually across time but also interact spatially, and the spatial and tempo…

2023

Transplayer: Timbre Style Transfer with Flexible Timbre Control

ICASSP 2023accepted

Music timbre style transfer aims at replacing the instrument timbre in a solo recording with another instrument, while preserving the musical content. Existing GAN-based methods can only achieve timbre style transfer between two given timbres. Inspired by the practice in voice conversion, we propose…

Cited by 0SourceScholar