← Search

Xiaohai Hu

3 accepted papers

2026

Multi-view Consistent Latent Action Learning for World Modeling and Control

ICML 2026poster

The scalability of world models is currently bottlenecked by the scarcity of action annotations. While self-supervised latent action learning offers a potential solution, existing single-view paradigms—relying on information bottlenecks or Vector Quantization (VQ)—often conflate superficial 2D pixel…

Cited by 0SourceScholar
2024

Context and Geometry Aware Voxel Transformer for Semantic Scene Completion

NeurIPS 2024spotlight

Vision-based Semantic Scene Completion (SSC) has gained much attention due to its widespread applications in various 3D perception tasks. Existing sparse-to-dense approaches typically employ shared context-independent queries across various input images, which fails to capture distinctions among the…

2024

Learned Slip-Detection-Severity Framework using Tactile Deformation Field Feedback for Robotic Manipulation

IROS 2024

Safely handling objects and avoiding slippage are fundamental challenges in robotic manipulation, yet traditional techniques often oversimplify the issue by treating slippage as a binary occurrence. Our research presents a framework that both identifies slip incidents and measures their severity. We

Cited by 8SourceScholar