← Search

Avisek Lahiri

3 accepted papers

2024

Directed Diffusion: Direct Control of Object Placement through Attention Guidance

AAAI 2024technical

Text-guided diffusion models such as DALLE-2, Imagen, and Stable Diffusion are able to generate an effectively endless variety of images given only a short text prompt describing the desired image content. In many cases the images are of very high quality. However, these models often struggle to com…

2021

LipSync3D: Data-Efficient Learning of Personalized 3D Talking Faces From Video Using Pose and Lighting Normalization

CVPR 2021poster

In this paper, we present a video-based learning framework for animating personalized 3D talking faces from audio. We introduce two training-time data normalizations that significantly improve data sample efficiency. First, we isolate and represent faces in a normalized space that decouples 3D geome…

Cited by 120PDFScholar
2020

Prior Guided GAN Based Semantic Inpainting

CVPR 2020poster

Contemporary deep learning based semantic inpainting can be approached from two directions. First, and the more explored, approach is to train an offline deep regression network over the masked pixels with an additional refinement by adversarial training. This approach requires a single feed-forward…

Cited by 124PDFScholar