← Search

Jiawei Hou

7 accepted papers

2026

DetAny4D: Detect Anything 4D Temporally in a Streaming RGB Video

CVPR 2026

Reliable 4D object detection, which refers to 3D object detection in streaming video, is crucial for perceiving and understanding the real world. Existing open-set 4D object detection methods typically make predictions on a frame-by-frame basis without modeling temporal consistency, or rely on compl

Cited by 0SourcecodeScholar
2026

LOG-Nav: Efficient Layout-Aware Object-Goal Navigation with Hierarchical Planning

AAAI 2026technical

We introduce LOG-Nav, an efficient layout-aware object-goal navigation approach designed for complex multi-room indoor environments. By planning hierarchically leveraging a global topologigal map with layout information and local imperative approach with detailed scene representation memory, LOG-Na

Cited by 0SourcePDFScholar
2025

Topo-Field: Topometric Mapping With Brain-Inspired Hierarchical Layout-Object-Position Fields

RA-L 2025

Mobile robots require comprehensive scene understanding to operate effectively in diverse environments, enriched with contextual information such as layouts, objects, and their relationships. Although advances like neural radiance fields (NeRFs) offer high-fidelity 3D reconstructions, they are compu

Cited by 3SourcecodeScholar
2024

FastOcc: Accelerating 3D Occupancy Prediction by Fusing the 2D Bird’s-Eye View and Perspective View

ICRA 2024poster

In autonomous driving, 3D occupancy prediction outputs voxel-wise status and semantic labels for more comprehensive understandings of 3D scenes compared with traditional perception tasks, such as 3D object detection and bird’s-eye view (BEV) semantic segmentation. Recent researchers have extensively…

Cited by 35SourceScholar
2024

TaMMa: Target-driven Multi-subscene Mobile Manipulation

CoRL 2024poster

For everyday service robotics, the ability to navigate back and forth based on tasks in multi-subscene environments and perform delicate manipulations is crucial and highly practical. While existing robotics primarily focus on complex tasks within a single scene or simple tasks across scalable scene…

Cited by 2SourceScholar
2023

FloorplanNet: Learning Topometric Floorplan Matching for Robot Localization

ICRA 2023poster

Given a building floorplan, humans can localize themselves by matching the observation of the environment with the floorplan using geometric, semantic, and topological clues. Inspired by this insight, this paper proposes a learning- based topometric robot localization method FloorplanNet, which impl…

Cited by 9SourcecodeScholar
2022

Multical: Spatiotemporal Calibration for Multiple IMUs, Cameras and LiDARs

IROS 2022poster

Spatiotemporal calibration of sensors, especially of those which do not share their fields of view, is becoming increasingly important in the fields of autonomous driving and robotics. This paper presents a general sensor calibration method, named Multical, that makes use of multiple planar calibrat…

Cited by 12SourceScholar