← Search

Danny Cohen-Or

7 accepted papers

2024

Be Yourself: Bounded Attention for Multi-Subject Text-to-Image Generation

ECCV 2024poster

"Text-to-image diffusion models have an unprecedented ability to generate diverse and high-quality images. However, they often struggle to faithfully capture the intended semantics of complex input prompts that include multiple subjects. Recently, numerous layout-to-image extensions have been introd…

Cited by 26SourcePDFScholar
2024

LCM-Lookahead for Encoder-based Text-to-Image Personalization

ECCV 2024poster

"Recent advancements in diffusion models have introduced fast sampling methods that can effectively produce high-quality images in just one or a few denoising steps. Interestingly, when these are distilled from existing diffusion models, they often maintain alignment with the original model, retaini…

2024

Lazy Diffusion Transformer for Interactive Image Editing

ECCV 2024poster

"We introduce a novel diffusion transformer, , that generates partial image updates efficiently. Our approach targets interactive image editing applications in which, starting from a blank canvas or an image, a user specifies a sequence of localized image modifications using binary masks and text pr…

Cited by 8SourcePDFScholar
2024

MyVLM: Personalizing VLMs for User-Specific Queries

ECCV 2024poster

"Recent large-scale vision-language models (VLMs) have demonstrated remarkable capabilities in understanding and generating textual descriptions for visual content. However, these models lack an understanding of user-specific concepts. In this work, we take a first step toward the personalization of…

Cited by 21SourcePDFScholar
2024

ReNoise: Real Image Inversion Through Iterative Noising

ECCV 2024poster

"Recent advancements in text-guided diffusion models have unlocked powerful image manipulation capabilities. However, applying these methods to real images necessitates the inversion of the images into the domain of the pretrained diffusion model. Achieving faithful inversion remains a challenge, pa…

Cited by 40SourcePDFScholar
2023

EmoSet: A Large-scale Visual Emotion Dataset with Rich Attributes

ICCV 2023poster

Visual Emotion Analysis (VEA) aims at predicting people's emotional responses to visual stimuli. This is a promising, yet challenging, task in affective computing, which has drawn increasing attention in recent years. Most of the existing work in this area focuses on feature design, while little att…

Cited by 52PDFScholar