← Search

Tian Yang

7 accepted papers

2025

Adjacent-view Transformers for Supervised Surround-view Depth Estimation

IROS 2025

Depth estimation has been widely studied and serves as the fundamental step of 3D perception for robotics and autonomous driving. Though significant progress has been made in monocular depth estimation in the past decades, these attempts are mainly conducted on the KITTI benchmark with only front-vi

Cited by 5SourcecodeScholar
2024

CoSLR: Contrastive Chinese Sign Language Recognition with prior knowledge And Multi-Tasks Joint Learning

ICASSP 2024accepted

Perceiving by computer vision, Sign Language Recognition (SLR) obtains the advantage of transforming the posture video into a sentence, compared with the methods of sensors to collect signals. However, learning representative features from a multimodal perspective is challenging. To this end, this s…

Cited by 0SourceScholar
2023

Conditional LS-GAN Based Skylight Polarization Image Restoration and Application in Meridian Localization

ICASSP 2023accepted

Skylight polarization images (SPIs) contain crucial spatial information that can be used for navigation purposes. Under most circumstances, the quality of the images becomes a major concern, especially when there is blocking between the perception equipment and the sky. This paper introduces a deep…

Cited by 0SourceScholar
2023

DyGait: Exploiting Dynamic Representations for High-performance Gait Recognition

ICCV 2023poster

Gait recognition is a biometric technology that recognizes the identity of humans through their walking patterns. Compared with other biometric technologies, gait recognition is more difficult to disguise and can be applied to the condition of long-distance without the cooperation of subjects. Thus,…

Cited by 48PDFScholar
2023

HFT: Lifting Perspective Representations via Hybrid Feature Transformation for BEV Perception

ICRA 2023poster

Restoring an accurate Bird's Eye View (BEV) map plays a crucial role in the perception of autonomous driving. The existing works of lifting representations from frontal view to BEV can be classified into two categories, i.e., Camera model-Based Feature Transformation (CBFT) and Camera model-Free Fea…

Cited by 11SourceScholar
2021

WebFace260M: A Benchmark Unveiling the Power of Million-Scale Deep Face Recognition

CVPR 2021poster

In this paper, we contribute a new million-scale face benchmark containing noisy 4M identities/260M faces (WebFace260M) and cleaned 2M identities/42M faces (WebFace42M) training data, as well as an elaborately designed time-constrained evaluation protocol. Firstly, we collect 4M name list and downlo…

Cited by 313PDFScholar