← Search

Jingjing Meng

8 accepted papers

2026

Pix2Key: Controllable Open-Vocabulary Retrieval with Semantic Decomposition and Self-Supervised Visual Dictionary Learning

ICML 2026poster

Composed image retrieval uses a reference image plus a natural-language edit to retrieve images that apply the requested change while preserving other relevant visual content. Classic fusion pipelines typically rely on supervised triplets and can lose fine-grained cues, while recent zero-shot approa…

Cited by 0SourceScholar
2024

Interaction-centric Spatio-Temporal Context Reasoning for Multi-Person Video HOI Recognition

ECCV 2024poster

"Understanding human-object interaction (HOI) in videos represents a fundamental yet intricate challenge in computer vision, requiring perception and reasoning across both spatial and temporal domains. Despite previous success of object detection and tracking, multi-person video HOI recognition stil…

2023

High Fidelity 3D Hand Shape Reconstruction via Scalable Graph Frequency Decomposition

CVPR 2023poster

Despite the impressive performance obtained by recent single-image hand modeling techniques, they lack the capability to capture sufficient details of the 3D hand mesh. This deficiency greatly limits their applications when high fidelity hand modeling is required, e.g., personalized hand modeling. T…

2023

Learning Attribute and Class-Specific Representation Duet for Fine-Grained Fashion Analysis

CVPR 2023poster

Fashion representation learning involves the analysis and understanding of various visual elements at different granularities and the interactions among them. Existing works often learn fine-grained fashion representations at the attribute-level without considering their relationships and inter-depe…

Cited by 12SourcePDFScholar
2019

Joint Representative Selection and Feature Learning: A Semi-Supervised Approach

CVPR 2019poster

In this paper, we propose a semi-supervised approach for representative selection, which finds a small set of representatives that can well summarize a large data collection. Given labeled source data and big unlabeled target data, we aim to find representatives in the target data, which can not onl…

Cited by 4PDFScholar
2016

From Keyframes to Key Objects: Video Summarization by Representative Object Proposal Selection

CVPR 2016poster

We propose to summarize a video into a few key objects by selecting representative object proposals generated from video frames. This representative selection problem is formulated as a sparse dictionary selection problem, i.e., choosing a few representatives object proposals to reconstruct the whol…

Cited by 136PDFScholar