← Search

Ziyin Zeng

4 accepted papers

2025

DeepLA-Net: Very Deep Local Aggregation Networks for Point Cloud Analysis

CVPR 2025poster

Due to the irregular and disordered data structure in 3D point clouds, prior works have focused on designing more sophisticated local representation methods to capture these complex local patterns. However, the recognition performance has saturated over the past few years, indicating that increasing…

2024

An Empirical Study of CLIP for Text-Based Person Search

AAAI 2024technical

Text-based Person Search (TBPS) aims to retrieve the person images using natural language descriptions. Recently, Contrastive Language Image Pretraining (CLIP), a universal large cross-modal vision-language pre-training model, has remarkably performed over various cross-modal downstream tasks due to…

2024

TALDS-Net: Task-Aware Adaptive Local Descriptors Selection for Few-Shot Image Classification

ICASSP 2024accepted

Few-shot image classification aims to classify images from unseen novel classes with few samples. Recent works demonstrate that deep local descriptors exhibit enhanced representational capabilities compared to image-level features. However, most existing methods solely rely on either employing all l…

Cited by 0SourceScholar
2023

An Empirical Study of Frame Selection for Text-to-Video Retrieval

EMNLP 2023long findings

Text-to-video retrieval (TVR) aims to find the most relevant video in a large video gallery given a query text. The intricate and abundant context of the video challenges the performance and efficiency of TVR. To handle the serialized video contexts, existing methods typically select a subset of fra…

Cited by 0SourceScholar