← Search

Boseung Jeong

6 accepted papers

2025

Learning Audio-guided Video Representation with Gated Attention for Video-Text Retrieval

CVPR 2025poster

Video-text retrieval, the task of retrieving videos based on a textual query or vice versa, is of paramount importance for video understanding and multimodal information retrieval. Recent methods in this area rely primarily on visual and textual features and often ignore audio, although it helps enh…

Cited by 0SourcePDFScholar
2024

Efficient and Versatile Robust Fine-Tuning of Zero-shot Models

ECCV 2024poster

"Large-scale image-text pre-trained models enable zero-shot classification and provide consistent accuracy across various data distributions. Nonetheless, optimizing these models in downstream tasks typically requires fine-tuning, which reduces generalization to out-of-distribution (OOD) data and de…

Cited by 3SourcePDFScholar
2024

PLOT: Text-based Person Search with Part Slot Attention for Corresponding Part Discovery

ECCV 2024poster

"Text-based person search, employing free-form text queries to identify individuals within a vast image collection, presents a unique challenge in aligning visual and textual representations, particularly at the human part level. Existing methods often struggle with part feature extraction and align…

Cited by 4SourcePDFScholar
2023

HIER: Metric Learning Beyond Class Labels via Hierarchical Regularization

CVPR 2023poster

Supervision for metric learning has long been given in the form of equivalence between human-labeled classes. Although this type of supervision has been a basis of metric learning for decades, we argue that it hinders further advances in the field. In this regard, we propose a new regularization met…

Cited by 20SourcePDFScholar
2023

Human Pose Estimation in Extremely Low-Light Conditions

CVPR 2023poster

We study human pose estimation in extremely low-light images. This task is challenging due to the difficulty of collecting real low-light images with accurate labels, and severely corrupted inputs that degrade prediction quality significantly. To address the first issue, we develop a dedicated camer…

2021

ASMR: Learning Attribute-Based Person Search With Adaptive Semantic Margin Regularizer

ICCV 2021poster

Attribute-based person search is the task of finding person images that are best matched with a set of text attributes given as query. The main challenge of this task is the large modality gap between attributes and images. To reduce the gap, we present a new loss for learning cross-modal embeddings…

Cited by 29PDFcodeScholar