← Search

Lele Cheng

3 accepted papers

2024

Decouple Content and Motion for Conditional Image-to-Video Generation

AAAI 2024technical

The goal of conditional image-to-video (cI2V) generation is to create a believable new video by beginning with the condition, i.e., one image and text. The previous cI2V generation methods conventionally perform in RGB pixel space, with limitations in modeling motion consistency and visual continuit…

Cited by 5SourcePDFScholar
2024

FashionERN: Enhance-and-Refine Network for Composed Fashion Image Retrieval

AAAI 2024technical

The goal of composed fashion image retrieval is to locate a target image based on a reference image and modified text. Recent methods utilize symmetric encoders (e.g., CLIP) pre-trained on large-scale non-fashion datasets. However, the input for this task exhibits an asymmetric nature, where the ref…

Cited by 5SourcePDFScholar
2020

Weakly Supervised Learning with Side Information for Noisy Labeled Images

ECCV 2020poster

In many real-world datasets, like WebVision, the performance of DNN based classier is often limited by the noisy labeled data. To tackle this problem, some image related side information, such as captions and tags, often reveal underlying relationships across images. In this paper, we present an eff…

Cited by 60SourcePDFScholar