← Search

Guorong Cai

5 accepted papers

2025

Depth Matters: Exploring Deep Interactions of RGB-D for Semantic Segmentation in Traffic Scenes

IROS 2025

RGB-D has gradually become a crucial data source for understanding complex scenes in assisted driving. However, existing studies have paid insufficient attention to the intrinsic spatial properties of depth maps. This oversight significantly impacts the attention representation, leading to predictio

Cited by 6SourceScholar
2025

Edge First: Edge-Guided Geometry for Superior 3D Roof Wireframe Reconstruction

ICASSP 2025accepted

Roof wireframe reconstruction has shown great success in 3D building reconstruction due to its lightweight nature and straightforward representation. However, previous methods consider all roof points, which result in edge redundancy and omissions. In this paper, we propose a novel and streamlined E…

Cited by 0SourceScholar
2025

EdgeDiff: Edge-aware Diffusion Network for Building Reconstruction from Point Clouds

CVPR 2025poster

Building reconstruction is a challenging problem at the intersection of computer vision, photogrammetry and computer graphics. 3D wireframe presents a compelling representation for building modeling through its compact structure. Existing wireframe reconstruction methods employing vertex detection a…

Cited by 0SourcePDFScholar
2025

Leveraging Depth and Language for Open-Vocabulary Domain-Generalized Semantic Segmentation

NeurIPS 2025poster

Open-Vocabulary semantic segmentation (OVSS) and domain generalization in semantic segmentation (DGSS) highlight a subtle complementarity that motivates Open-Vocabulary Domain-Generalized Semantic Segmentation (OV-DGSS). OV-DGSS aims to generate pixel-level masks for unseen categories while maintain…

Cited by 0SourcecodeScholar
2025

Stronger, Steadier & Superior: Geometric Consistency in Depth VFM Forges Domain Generalized Semantic Segmentation

ICCV 2025poster

Vision Foundation Models (VFMs) have delivered remarkable performance in Domain Generalized Semantic Segmentation (DGSS). However, recent methods often overlook the fact that visual cues are susceptible, whereas the underlying geometry remains stable, rendering depth information more robust. In this…