← Search

Antonio Alliegro

5 accepted papers

2024

A Backpack Full of Skills: Egocentric Video Understanding with Diverse Task Perspectives

CVPR 2024poster

Human comprehension of a video stream is naturally broad: in a few instants we are able to understand what is happening the relevance and relationship of objects and forecast what will follow in the near future everything all at once. We believe that - to effectively transfer such an holistic percep…

2024

MeshGPT: Generating Triangle Meshes with Decoder-Only Transformers

CVPR 2024highlight

We introduce MeshGPT a new approach for generating triangle meshes that reflects the compactness typical of artist-created meshes in contrast to dense triangle meshes extracted by iso-surfacing methods from neural fields. Inspired by recent advances in powerful large language models we adopt a seque…

Cited by 124SourcePDFScholar
2022

3DOS: Towards 3D Open Set Learning - Benchmarking and Understanding Semantic Novelty Detection on Point Clouds

NeurIPS 2022accept

In recent years there has been significant progress in the field of 3D learning on classification, detection and segmentation problems. The vast majority of the existing studies focus on canonical closed-set conditions, neglecting the intrinsic open nature of the real-world. This limits the abilitie…

2022

End-to-End Learning to Grasp via Sampling From Object Point Clouds

RA-L 2022

The ability to grasp objects is an essential skill that enables many robotic manipulation tasks. Recent works have studied point cloud-based methods for object grasping by starting from simulated datasets and have shown promising performance in real-world scenarios. Nevertheless, many of them still

Cited by 35SourcecodeScholar
2021

Denoise and Contrast for Category Agnostic Shape Completion

CVPR 2021poster

In this paper, we present a deep learning model that exploits the power of self-supervision to perform 3D point cloud completion, estimating the missing part and a context region around it. Local and global information are encoded in a combined embedding. A denoising pretext task provides the networ…

Cited by 46PDFcodeScholar