← Search

Xuewei Tony Qi

2 accepted papers

2025

ET-Former: Efficient Triplane Deformable Attention for 3D Semantic Scene Completion From Monocular Camera

IROS 2025

We introduce ET-Former, a novel end-to-end algorithm for semantic scene completion using a single monocular camera. Our approach generates a semantic occupancy map from single RGB observation while simultaneously providing uncertainty estimates for semantic predictions. By designing a triplane-based

Cited by 3SourcecodeScholar
2024

VLPG-Nav: Object Navigation Using Visual Language Pose Graph and Object Localization Probability Maps

IROS 2024poster

We present VLPG-Nav, a visual language navigation method for guiding robots to specified objects within household scenes. Unlike existing methods primarily focused on navigating the robot toward objects, our approach considers the additional challenge of centering the object within the robot’s camer…

Cited by 1SourceScholar