← Search

Shengchao Zhou

4 accepted papers

2026

ASSIST-3D: Adapted Scene Synthesis for Class-Agnostic 3D Instance Segmentation

AAAI 2026technical

Class-agnostic 3D instance segmentation tackles the challenging task of segmenting all object instances, including previously unseen ones, without semantic class reliance. Current methods struggle with generalization due to the scarce annotated 3D scene data or noisy 2D segmentations. While syntheti

Cited by 0SourcePDFScholar
2026

Learning to Reason in 4D: Dynamic Spatial Understanding for Vision Language Models

CVPR 2026

Vision-language models (VLM) excel at general understanding yet remain weak at dynamic spatial reasoning (DSR), i.e., reasoning about the evolvement of object geometry and relationship in 3D space over time, largely due to the scarcity of scalable 4D-aware training resources. To bridge this gap acro

Cited by 0SourcecodeScholar
2023

Robust Feature Rectification of Pretrained Vision Models for Object Recognition

AAAI 2023technical

Pretrained vision models for object recognition often suffer a dramatic performance drop with degradations unseen during training. In this work, we propose a RObust FEature Rectification module (ROFER) to improve the performance of pretrained models against degradations. Specifically, ROFER first es…

Cited by 0SourcePDFScholar
2023

UniDistill: A Universal Cross-Modality Knowledge Distillation Framework for 3D Object Detection in Bird's-Eye View

CVPR 2023highlight

In the field of 3D object detection for autonomous driving, the sensor portfolio including multi-modality and single-modality is diverse and complex. Since the multi-modal methods have system complexity while the accuracy of single-modal ones is relatively low, how to make a tradeoff between them is…