← Search

Tiesong Zhao

4 accepted papers

2026

Syntactic Structure-Guided Visual Grounding with Subject-Centric Feature Enhancement and Verification

IJCAI 2026

Visual grounding aims to localize target objects based on natural language descriptions, and the core challenge lies in the cross-modal gap, which is partly caused by the significant differences in semantic structure between language and vision. Existing methods typically rely on holistic sentence-l

Cited by 0Scholar
2025

Keep the Balance: A Parameter-Efficient Symmetrical Framework for RGB+X Semantic Segmentation

CVPR 2025poster

Multimodal semantic segmentation is a critical challenge in computer vision, with early methods suffering from high computational costs and limited transferability due to full fine-tuning of RGB-based pre-trained parameters. Recent studies, while leveraging additional modalities as supplementary pro…

Cited by 0SourcePDFScholar