← Search

Sara Vicente

9 accepted papers

2025

PlaceIt3D: Language-Guided Object Placement in Real 3D Scenes

ICCV 2025poster

We introduce the task of Language-Guided Object Placement in Real 3D Scenes. Given a 3D reconstructed point-cloud scene, a 3D asset, and a natural-language instruction, the goal is to place the asset so that the instruction is satisfied. The task demands tackling four intertwined challenges: (a) one…

Cited by 0SourcePDFScholar
2024

AirPlanes: Accurate Plane Estimation via 3D-Consistent Embeddings

CVPR 2024poster

Extracting planes from a 3D scene is useful for downstream tasks in robotics and augmented reality. In this paper we tackle the problem of estimating the planar surfaces in a scene from posed images. Our first finding is that a surprisingly competitive baseline results from combining popular cluster…

Cited by 1SourcePDFScholar
2024

DoubleTake: Geometry Guided Depth Estimation

ECCV 2024poster

"Estimating depth from a sequence of posed RGB images is a fundamental computer vision task, with applications in augmented reality, path planning etc. Prior work typically makes use of previous frames in a multi view stereo framework, relying on matching textures in a local neighborhood. In contras…

Cited by 1SourcePDFScholar
2023

Removing Objects From Neural Radiance Fields

CVPR 2023poster

Neural Radiance Fields (NeRFs) are emerging as a ubiquitous scene representation that allows for novel view synthesis. Increasingly, NeRFs will be shareable with other people. Before sharing a NeRF, though, it might be desirable to remove personal information or unsightly objects. Such removal is no…

Cited by 71SourcePDFScholar
2023

Virtual Occlusions Through Implicit Depth

CVPR 2023poster

For augmented reality (AR), it is important that virtual assets appear to 'sit among' real world objects. The virtual element should variously occlude and be occluded by real matter, based on a plausible depth ordering. This occlusion should be consistent over time as the viewer's camera moves. Unfo…

2022

Map-Free Visual Relocalization: Metric Pose Relative to a Single Image

ECCV 2022poster

"Can we relocalize in a scene represented by a single reference image? Standard visual relocalization requires hundreds of images and scale calibration to build a scene-specific 3D map. In contrast, we propose Map-free Relocalization, i.e., using only one photo of a scene to enable instant, metric s…

2020

The GAN That Warped: Semantic Attribute Editing With Unpaired Data

CVPR 2020poster

Deep neural networks have recently been used to edit images with great success, in particular for faces. However, they are often limited to only being able to work at a restricted range of resolutions. Many methods are so flexible that face edits can often result in an unwanted loss of identity. Thi…

Cited by 27PDFScholar
2018

Structured Uncertainty Prediction Networks

CVPR 2018poster

This paper is the first work to propose a network to predict a structured uncertainty distribution for a synthesized image. Previous approaches have been mostly limited to predicting diagonal covariance matrices. Our novel model learns to predict a full Gaussian covariance matrix for each reconstruc…

Cited by 81SourcePDFScholar