← Search

Axel Barroso-Laguna

7 accepted papers

2026

A Scene is Worth a Thousand Features: Feed-Forward Camera Localization from a Collection of Image Features

ICLR 2026poster

Visually localizing an image, i.e., estimating its camera pose, requires building a scene representation that serves as a visual map. The representation we choose has direct consequences towards the practicability of our system. Even when starting from mapping images with known camera poses, state-o…

Cited by 0SourceScholar
2025

ACE-G: Improving Generalization of Scene Coordinate Regression Through Query Pre-Training

ICCV 2025poster

Scene coordinate regression (SCR) has established itself as a promising learning-based approach to visual relocalization. After mere minutes of scene-specific training, SCR models estimate camera poses of query images with high accuracy. Still, SCR methods fall short of the generalization capabiliti…

Cited by 0SourcePDFScholar
2025

Scene Coordinate Reconstruction Priors

ICCV 2025poster

Scene coordinate regression (SCR) models have proven to be powerful implicit scene representations for 3D vision, enabling visual relocalization and structure-from-motion. SCR models are trained specifically for one scene. If training images imply insufficient multi-view constraints to recover the s…

Cited by 0SourcePDFScholar
2024

Matching 2D Images in 3D: Metric Relative Pose from Metric Correspondences

CVPR 2024poster

Given two images we can estimate the relative camera pose between them by establishing image-to-image correspondences. Usually correspondences are 2D-to-2D and the pose we estimate is defined only up to scale. Some applications aiming at instant augmented reality anywhere require scale-metric pose e…

2023

Two-View Geometry Scoring Without Correspondences

CVPR 2023poster

Camera pose estimation for two-view geometry traditionally relies on RANSAC. Normally, a multitude of image correspondences leads to a pool of proposed hypotheses, which are then scored to find a winning model. The inlier count is generally regarded as a reliable indicator of "consensus". We examine…

2019

Key.Net: Keypoint Detection by Handcrafted and Learned CNN Filters

ICCV 2019poster

We introduce a novel approach for keypoint detection task that combines handcrafted and learned CNN filters within a shallow multi-scale architecture. Handcrafted filters provide anchor structures for learned filters, which localize, score and rank repeatable features. Scale-space representation is…

Cited by 364PDFcodeScholar