← Search

Shiqi Zhao

5 accepted papers

2026

Personality-guided Public-Private Domain Disentangled Hypergraph-Former Network for Multimodal Depression Detection

AAAI 2026technical

Depression represents a global mental health challenge requiring efficient and reliable automated detection methods. Current Transformer- or Graph Neural Networks (GNNs)-based multimodal depression detection methods face significant challenges in modeling individual differences and cross-modal tempo

Cited by 0SourcePDFScholar
2024

Multi-Level Contrastive Learning For Hybrid Cross-Modal Retrieval

ICASSP 2024accepted

Hybrid image retrieval is a significant task for a wide range of applications. In this scenario, the hybrid query for searching images consists of a reference image and a text modifier. The reference image provides a vital visual context and displays some semantic details, while the text modifier sp…

Cited by 0SourceScholar
2024

Unsupervised Continual Learning of Image Representation Via Rememory-Based Simsiam

ICASSP 2024accepted

Unsupervised continual learning (UCL) of image representation has garnered attention due to practical need. However, recent UCL methods focus on mitigating the catastrophic forgetting with a replay buffer (i.e., rehearsal-based strategy), which needs much extra storage. To overcome this drawback, we…

Cited by 0SourceScholar
2023

MUI-TARE: Cooperative Multi-Agent Exploration With Unknown Initial Position

RA-L 2023

Multi-agent exploration of a bounded 3D environment with the unknown initial poses of agents is a challenging problem. It requires both quickly exploring the environments and robustly merging the sub-maps built by the agents. Most existing exploration strategies directly merge two sub-maps built by

Cited by 25SourceScholar
2023

SphereVLAD++: Attention-Based and Signal-Enhanced Viewpoint Invariant Descriptor

RA-L 2023

LiDAR-based localization approach is a fundamental module for large-scale navigation tasks, such as last-mile delivery and autonomous driving, and localization robustness highly relies on viewpoints and 3D feature extraction. Our previous work provides a viewpoint-invariant descriptor to deal with v

Cited by 26SourceScholar