← Search

Zijun Xu

8 accepted papers

2026

CMoE: Contrastive Mixture of Experts for Motion Control and Terrain Adaptation of Humanoid Robots

ICRA 2026poster

For effective deployment in real-world environments, humanoid robots must autonomously navigate a diverse range of complex terrains with abrupt transitions. While the Vanilla mixture of experts (MoE) framework is theoretically capable of modeling diverse terrain features, in practice, the gating net…

2026

DiffuView: Multi-View Diffusion Pretraining for 3D Aware Robotic Manipulation

CVPR 2026

Robotic manipulation from visual observations remains challenging due to the lack of 3D consistent representations that can generalize across diverse viewpoints and sensor configurations. Existing approaches often rely on masked autoencoders or neural scene representations, which fail to capture cro

Cited by 0SourceScholar
2026

Drive in Corridors: Enhancing the Safety of End-To-End Autonomous Driving Via Corridor Learning and Planning

ICRA 2026poster

Safety remains one of the most critical challenges in autonomous driving systems. In recent years, the end-to-end driving has shown great promise in advancing vehicle autonomy in a scalable manner. However, existing approaches often face safety risks due to the lack of explicit behavior constraints.…

2026

Rhythm: Learning Interactive Whole-Body Control for Dual Humanoids

RSS 2026poster

Realizing interactive whole-body control for multi-humanoid systems is critical for unlocking complex collaborative capabilities in shared environments. Although recent advancements have significantly enhanced the agility of individual robots, bridging the gap to physically coupled multi-humanoid in…

Cited by 1SourceScholar
2025

Drive in Corridors: Enhancing the Safety of End-to-End Autonomous Driving via Corridor Learning and Planning

RA-L 2025

Safety remains one of the most critical challenges in autonomous driving systems. In recent years, the end-to-end driving has shown great promise in advancing vehicle autonomy in a scalable manner. However, existing approaches often face safety risks due to the lack of explicit behavior constraints.

Cited by 3SourcecodeScholar
2025

FriendsQA: A New Large-Scale Deep Video Understanding Dataset with Fine-grained Topic Categorization for Story Videos

AAAI 2025technical

Video question answering (VideoQA) aims to answer natural language questions according to the given videos. Although existing models perform well in the factoid VideoQA task, they still face challenges in deep video understanding (DVU) task, which focuses on story videos. Compared to factoid videos,…

2025

HGS-Planner: Hierarchical Planning Framework for Active Scene Reconstruction Using 3D Gaussian Splatting

ICRA 2025

In complex missions such as search and rescue, robots must make intelligent decisions in unknown environments, relying on their ability to perceive and understand their surroundings. High-quality and real-time reconstruction enhances situational awareness and is crucial for intelligent robotics. Tra

Cited by 19SourceScholar
2025

PS-CC: Planning by Simulation With Continuous Action Space Optimization and Adaptive Cost Learning for Human-Like Autonomous Driving

RA-L 2025

Ensuring safe and human-like decision-making is a critical component of autonomous vehicle decision systems. However, conventional approaches often simplify the action space into discrete domains for policy formulation and rely on handcrafted cost functions for policy evaluation, which limits their

Cited by 0SourceScholar