← Search

Emanuele Aiello

4 accepted papers

2025

DreamCache: Finetuning-Free Lightweight Personalized Image Generation via Feature Caching

CVPR 2025poster

Personalized image generation requires text-to-image generative models that capture the core features of a reference subject to allow for controlled generation across different contexts. Existing methods face challenges due to complex training requirements, high inference costs, limited flexibility,…

2024

Jointly Training Large Autoregressive Multimodal Models

ICLR 2024poster

In recent years, advances in the large-scale pretraining of language and text-to-image models have revolutionized the field of machine learning. Yet, integrating these two modalities into a single, robust model capable of generating seamless multimodal outputs remains a significant challenge. To add…

Cited by 33SourcePDFScholar
2024

MotionCraft: Physics-Based Zero-Shot Video Generation

NeurIPS 2024poster

Generating videos with realistic and physically plausible motion is one of the main recent challenges in computer vision. While diffusion models are achieving compelling results in image generation, video diffusion models are limited by heavy training and huge models, resulting in videos that are s…

2022

Cross-modal Learning for Image-Guided Point Cloud Shape Completion

NeurIPS 2022accept

In this paper we explore the recent topic of point cloud completion, guided by an auxiliary image. We show how it is possible to effectively combine the information from the two modalities in a localized latent space, thus avoiding the need for complex point cloud reconstruction methods from single…