← Search

Marco Manfredi

3 accepted papers

2026

Fast SceneScript: Fast and Accurate Language-Based 3D Scene Understanding via Multi-Token Prediction

CVPR 2026

Recent perception-generalist approaches based on language models have achieved state-of-the-art results across diverse tasks, including 3D scene layout estimation and 3D object detection, via unified architecture and interface. However, these approaches rely on autoregressive next-token prediction,

Cited by 0SourceScholar
2022

Visual Cross-View Metric Localization with Dense Uncertainty Estimates

ECCV 2022poster

"This work addresses visual cross-view metric localization for outdoor robotics. Given a ground-level color image and a satellite patch that contains the local surroundings, the task is to identify the location of the ground camera within the satellite patch. Related work addressed this task for ran…

2021

Cross-View Matching for Vehicle Localization by Learning Geographically Local Representations

RA-L 2021

Cross-view matching aims to learn a shared image representation between ground-level images and satellite or aerial images at the same locations. In robotic vehicles, matching a camera image to a database of geo-referenced aerial imagery can serve as a method for self-localization. However, existing

Cited by 27SourceScholar