← Search

Dingzeyu Li

3 accepted papers

2024

Concept Weaver: Enabling Multi-Concept Fusion in Text-to-Image Models

CVPR 2024poster

While there has been significant progress in customizing text-to-image generation models generating images that combine multiple personalized concepts remains challenging. In this work we introduce Concept Weaver a method for composing customized text-to-image diffusion models at inference time. Spe…

Cited by 12SourcePDFScholar
2022

Audio-Driven Neural Gesture Reenactment With Video Motion Graphs

CVPR 2022poster

Human speech is often accompanied by body gestures including arm and hand gestures. We present a method that reenacts a high-quality video with gestures matching a target speech audio. The key idea of our method is to split and re-assemble clips from a reference video through a novel video motion gr…

Cited by 19PDFcodeScholar
2020

Unified Multisensory Perception: Weakly-Supervised Audio-Visual Video Parsing

ECCV 2020poster

In this paper, we introduce a new problem, named audio-visual video parsing, which aims to parse a video into temporal event segments and label them as either audible, visible, or both. Such a problem is essential for a complete understanding of the scene depicted inside a video. To facilitate explo…