← Search

Nicolas Dufour

6 accepted papers

2026

MIRO: MultI-Reward cOnditioned pretraining improves T2I quality and efficiency

ICML 2026poster

The default paradigm of post-training text-to-image generators includes post-hoc selection of generated images, and subsequent training with one reward model to align the generator to the reward, typically user preference. This discards informative data as well as optimizes only for a single reward,…

Cited by 0SourceScholar
2025

Around the World in 80 Timesteps: A Generative Approach to Global Visual Geolocation

CVPR 2025poster

Global visual geolocation predicts where an image was captured on Earth. Since images vary in how precisely they can be localized, this task inherently involves a significant degree of ambiguity. However, existing approaches are deterministic and overlook this aspect. In this paper, we aim to close…

2024

Don't Drop Your Samples! Coherence-Aware Training Benefits Conditional Diffusion

CVPR 2024highlight

Conditional diffusion models are powerful generative models that can leverage various types of conditional information such as class labels segmentation masks or text captions. However in many real-world scenarios conditional information may be noisy or unreliable due to human annotation errors or w…

Cited by 3SourcePDFScholar
2024

E.T. the Exceptional Trajectory: Text-to-camera-trajectory generation with character awareness

ECCV 2024poster

"Stories and emotions in movies emerge through the effect of well-thought-out directing decisions, in particular camera placement and movement over time. Crafting compelling camera trajectories remains a complex iterative process, even for skilful artists. To tackle this, in this paper, we propose a…

Cited by 3SourcePDFScholar
2024

OpenStreetView-5M: The Many Roads to Global Visual Geolocation

CVPR 2024poster

Determining the location of an image anywhere on Earth is a complex visual task which makes it particularly relevant for evaluating computer vision algorithms. Determining the location of an image anywhere on Earth is a complex visual task which makes it particularly relevant for evaluating computer…

2022

SCAM! Transferring Humans between Images with Semantic Cross Attention Modulation

ECCV 2022poster

"A large body of recent work targets semantically conditioned image generation. Most such methods focus on the narrower task of pose transfer and ignore the more challenging task of subject transfer that consists in not only transferring the pose but also the appearance and background. In this work,…