← Search

Moustafa Meshry

4 accepted papers

2024

Rethinking Video-Text Understanding: Retrieval from Counterfactually Augmented Data

ECCV 2024poster

"Recent video-text foundation models have demonstrated strong performance on a wide variety of downstream video understanding tasks. Can these video-text models genuinely understand the contents of natural videos? Standard video-text evaluations could be misleading as many questions can be inferred…

Cited by 2SourcePDFScholar
2021

Learned Spatial Representations for Few-Shot Talking-Head Synthesis

ICCV 2021poster

We propose a novel approach for few-shot talking-head synthesis. While recent works in neural talking heads have produced promising results, they can still produce images that do not preserve the identity of the subject in source images. We posit this is a result of the entangled representation of e…

Cited by 49PDFScholar
2021

StEP: Style-Based Encoder Pre-Training for Multi-Modal Image Synthesis

CVPR 2021poster

We propose a novel approach for multi-modal Image-to-image (I2I) translation. To tackle the one-to-many relationship between input and output domains, previous works use complex training objectives to learn a latent embedding, jointly with the generator, that models the variability of the output dom…

Cited by 11PDFScholar