← Search

Hongda Mao

5 accepted papers

2025

Efficient Visual Place Recognition Through Multimodal Semantic Knowledge Integration

ICCV 2025poster

Visual place recognition is crucial for autonomous navigation and robotic mapping. Current methods struggle with perceptual aliasing and computational inefficiency. We present SemVPR, a novel approach integrating multimodal semantic knowledge into VPR. By leveraging a pre-trained vision-language mod…

Cited by 0SourcePDFScholar
2025

Head2Body: Body Pose Generation from Multi-sensory Head-mounted Inputs

ICCV 2025poster

Generating body pose from head-mounted, egocentric inputs is essential for immersive VR/AR and assistive technologies, as it supports more natural interactions. However, the task is challenging due to limited visibility of body parts in first-person views and the sparseness of sensory data, with onl…

Cited by 0SourcePDFScholar
2023

Augmentation Robust Self-Supervised Learning for Human Activity Recognition

ICASSP 2023accepted

Human Activity Recognition (HAR) is widely applied on wearable devices in our daily lives. However, acquiring high-quality wearable sensor data set with ground-truths is challenging due to the high cost in collecting data and necessity of domain experts. In order to achieve generalization from limit…

Cited by 0SourceScholar
2020

Bridging Mixture Density Networks with Meta-Learning for Automatic Speaker Identification

ICASSP 2020accepted

Speaker identification answers the fundamental question "Who is speaking" The identification technology enables various downstream applications to provide a personalized experience. Both the prevalent i-vector based solutions and the state-of-the-art deep learning solutions usually treat all users e…

Cited by 0SourceScholar