← Search

Luigi Riz

4 accepted papers

2026

Efficient Encoder-Free Fourier-based 3D Large Multimodal Model

CVPR 2026

Large Multimodal Models (LMMs) that process 3D data typically rely on heavy, pretrained visual encoders to extract geometric features. While recent 2D LMMs have begun to eliminate such encoders for efficiency and scalability, extending this paradigm to 3D remains challenging due to the unordered and

Cited by 0SourceScholar
2024

Geometrically-driven Aggregation for Zero-shot 3D Point Cloud Understanding

CVPR 2024highlight

Zero-shot 3D point cloud understanding can be achieved via 2D Vision-Language Models (VLMs). Existing strategies directly map VLM representations from 2D pixels of rendered or captured views to 3D points overlooking the inherent and expressible point cloud geometric structure. Geometrically similar…

2023

Novel Class Discovery for 3D Point Cloud Semantic Segmentation

CVPR 2023poster

Novel class discovery (NCD) for semantic segmentation is the task of learning a model that can segment unlabelled (novel) classes using only the supervision from labelled (base) classes. This problem has recently been pioneered for 2D image data, but no work exists for 3D point cloud data. In fact,…