← Search

Fabio Remondino

4 accepted papers

2025

Renderworld: World Model with Self-Supervised 3D Label

ICRA 2025

End-to-end autonomous driving with vision-only is not only more cost-effective compared to LiDAR-vision fusion but also more reliable than traditional methods. To achieve a economical and robust purely visual autonomous driving system, we propose RenderWorld, a vision-only end-to-end autonomous driv

Cited by 47SourceScholar
2024

Towards Enhanced Human Activity Recognition for Real-World Human-Robot Collaboration

ICRA 2024poster

This research contributes to the field of Human-Robot Collaboration (HRC) within dynamic and unstructured environments by extending the previously proposed Fuzzy State-Long Short-Term Memory (FS-LSTM) architecture to handle the uncertainty and irregularity inherent in real-world sensor data. Recogni…

Cited by 3SourceScholar
2020

Image-to-Voxel Model Translation for 3D Scene Reconstruction and Segmentation

ECCV 2020poster

Objects class, depth, and shape are instantly reconstructed by a human looking at a 2D image. While modern deep models solve each of these challenging tasks separately, they struggle to perform simultaneous scene 3D reconstruction and segmentation. We propose a single shot image-to-semantic voxel mo…

Cited by 23SourcePDFScholar
2019

The Point Where Reality Meets Fantasy: Mixed Adversarial Generators for Image Splice Detection

NeurIPS 2019poster

Modern photo editing tools allow creating realistic manipulated images easily. While fake images can be quickly generated, learning models for their detection is challenging due to the high variety of tampering artifacts and the lack of large labeled datasets of manipulated images. In this paper, we…