← Search

Gabriela Csurka

15 accepted papers

2025

Gaussian Splatting Feature Fields for (Privacy-Preserving) Visual Localization

CVPR 2025poster

Visual localization is the task of estimating a camera pose in a known environment. In this paper, we utilize 3D Gaussian Splatting (3DGS)-based representations for accurate and privacy-preserving visual localization. We propose Gaussian Splatting Feature Fields (GSFFs), a scene representation for v…

Cited by 0SourcePDFScholar
2025

MUSt3R: Multi-view Network for Stereo 3D Reconstruction

CVPR 2025highlight

DUSt3R introduced a novel paradigm in geometric computer vision by proposing a model that can provide dense and unconstrained Stereo 3D Reconstruction of arbitrary image collections with no prior information about camera calibration nor viewpoint poses. Under the hood, however, DUSt3R processes imag…

2025

PanSt3R: Multi-view Consistent Panoptic Segmentation

ICCV 2025poster

Panoptic segmentation in 3D is a fundamental problem in scene understanding. Existing approaches typically rely on costly test-time optimizations (often based on NeRF) to consolidate 2D predictions of off-the-shelf panoptic segmentation methods into 3D. Instead, in this work, we propose a unified an…

Cited by 0SourcePDFScholar
2024

SHiNe: Semantic Hierarchy Nexus for Open-vocabulary Object Detection

CVPR 2024highlight

Open-vocabulary object detection (OvOD) has transformed detection into a language-guided task empowering users to freely define their class vocabularies of interest during inference. However our initial investigation indicates that existing OvOD detectors exhibit significant variability when dealing…

2024

Weatherproofing Retrieval for Localization with Generative AI and Geometric Consistency

ICLR 2024poster

State-of-the-art visual localization approaches generally rely on a first image retrieval step whose role is crucial. Yet, retrieval often struggles when facing varying conditions, due to e.g. weather or time of day, with dramatic consequences on the visual localization accuracy. In this paper, we i…

Cited by 0SourcePDFScholar
2023

CroCo v2: Improved Cross-view Completion Pre-training for Stereo Matching and Optical Flow

ICCV 2023poster

Despite impressive performance for high-level downstream tasks, self-supervised pre-training methods have not yet fully delivered on dense geometric vision tasks such as stereo matching or optical flow. The application of self-supervised concepts, such as instance discrimination or masked image mode…

Cited by 100PDFcodeScholar
2023

SegLoc: Learning Segmentation-Based Representations for Privacy-Preserving Visual Localization

CVPR 2023poster

Inspired by properties of semantic segmentation, in this paper we investigate how to leverage robust image segmentation in the context of privacy-preserving visual localization. We propose a new localization framework, SegLoc, that leverages image segmentation to create robust, compact, and privacy-…

Cited by 17SourcePDFScholar
2022

ARTEMIS: Attention-based Retrieval with Text-Explicit Matching and Implicit Similarity

ICLR 2022poster

An intuitive way to search for images is to use queries composed of an example image and a complementary text. While the first provides rich and implicit context for the search, the latter explicitly calls for new traits, or specifies how some elements of the example image should be changed to retri…

2022

CroCo: Self-Supervised Pre-training for 3D Vision Tasks by Cross-View Completion

NeurIPS 2022accept

Masked Image Modeling (MIM) has recently been established as a potent pre-training paradigm. A pretext task is constructed by masking patches in an input image, and this masked content is then predicted by a neural network using visible patches as sole input. This pre-training leads to state-of-the-…

2022

Deep Visual Geo-Localization Benchmark

CVPR 2022oral

In this paper, we propose a new open-source benchmarking framework for Visual Geo-localization (VG) that allows to build, train, and test a wide range of commonly used architectures, with the flexibility to change individual components of a geo-localization pipeline. The purpose of this framework is…

Cited by 103PDFcodeScholar
2022

On the Road to Online Adaptation for Semantic Image Segmentation

CVPR 2022poster

We propose a new problem formulation and a corresponding evaluation framework to advance research on unsupervised domain adaptation for semantic image segmentation. The overall goal is fostering the development of adaptive learning systems that will continuously learn, without supervision, in ever-c…

Cited by 35PDFcodeScholar
2021

Large-Scale Localization Datasets in Crowded Indoor Spaces

CVPR 2021poster

Estimating the precise location of a camera using visual localization enables interesting applications such as augmented reality or robot navigation. This is particularly useful in indoor environments where other localization technologies, such as GNSS, fail. Indoor spaces impose interesting challen…

Cited by 50PDFcodeScholar
2019

Fine-Grained Action Retrieval Through Multiple Parts-of-Speech Embeddings

ICCV 2019poster

We address the problem of cross-modal fine-grained action retrieval between text and video. Cross-modal retrieval is commonly achieved through learning a shared embedding space, that can indifferently embed modalities. In this paper, we propose to enrich the embedding by disentangling parts-of-speec…

Cited by 184PDFScholar
2019

Visual Localization by Learning Objects-Of-Interest Dense Match Regression

CVPR 2019poster

We introduce a novel CNN-based approach for visual localization from a single RGB image that relies on densely matching a set of Objects-of-Interest (OOIs). In this paper, we focus on planar objects which are highly descriptive in an environment, such as paintings in museums or logos and storefronts…

Cited by 53PDFScholar