← Search

Sitao Zhang

3 accepted papers

2025

Efficient Visual Place Recognition Through Multimodal Semantic Knowledge Integration

ICCV 2025poster

Visual place recognition is crucial for autonomous navigation and robotic mapping. Current methods struggle with perceptual aliasing and computational inefficiency. We present SemVPR, a novel approach integrating multimodal semantic knowledge into VPR. By leveraging a pre-trained vision-language mod…

Cited by 0SourcePDFScholar
2025

S2S2: Semantic Stacking for Robust Semantic Segmentation in Medical Imaging

AAAI 2025technical

Robustness and generalizability in medical image segmentation are often hindered by scarcity and limited diversity of training data, which stands in contrast to the variability encountered during inference. While conventional strategies---such as domain-specific augmentation, specialized architectur…

2023

Learning Emotion Representations From Verbal and Nonverbal Communication

CVPR 2023poster

Emotion understanding is an essential but highly challenging component of artificial general intelligence. The absence of extensive annotated datasets has significantly impeded advancements in this field. We present EmotionCLIP, the first pre-training paradigm to extract visual emotion representatio…