← Search

Tianyi Shang

4 accepted papers

2026

FourierPlace: A Vision-Language Localization Framework Based on Frequency Domain Representations

RA-L 2026

Language-guided localization within 3D environments continues to pose a significant challenge for autonomous systems, primarily due to the need for precise alignment between sparse point cloud data and inherently ambiguous natural language descriptions. To address this, we present a novel vision-lan

Cited by 0SourcecodeScholar
2025

Bridging Text and Vision: A Multi-View Text-Vision Registration Approach for Cross-Modal Place Recognition

IROS 2025

Mobile robots necessitate advanced natural language understanding capabilities to accurately identify locations and perform tasks such as package delivery. However, traditional visual place recognition (VPR) methods rely solely on single-view visual information and cannot interpret human language de

Cited by 6SourcecodeScholar
2025

MambaPlace: Text-to-Point-Cloud Cross-Modal Place Recognition with Attention Mamba Mechanisms

IROS 2025

Vision-Language Place Recognition (VLPR) enhances robot localization performance by incorporating natural language descriptions from images. By utilizing language information, VLPR directs robot place matching, overcoming the constraint of solely depending on vision. However, general multimodal info

Cited by 8SourcecodeScholar