← Search

Diego Gomez

3 accepted papers

2025

Escaping Plato's Cave: Towards the Alignment of 3D and Text Latent Spaces

CVPR 2025poster

Recent works have shown that, when trained at scale, uni-modal 2D vision and text encoders converge to learned features that share remarkable structural properties, despite arising from different representations. However, the role of 3D encoders with respect to other modalities remains unexplored. F…

Cited by 0SourcePDFScholar
2025

ZeroKey: Point-Level Reasoning and Zero-Shot 3D Keypoint Detection from Large Language Models

ICCV 2025poster

We propose a novel zero-shot approach for keypoint detection on 3D shapes. Point-level reasoning on visual data is challenging as it requires precise localization capability, posing problems even for powerful models like DINO or CLIP. Traditional methods for 3D keypoint detection rely heavily on ann…

Cited by 0SourcePDFScholar