2025
SceneSplat: Gaussian Splatting-based Scene Understanding with Vision-Language Pretraining
ICCV 2025poster
Recognizing arbitrary or previously unseen categories is essential for comprehensive real-world 3D scene understanding. Currently, all existing methods rely on 2D or textual modalities during training, or together at inference. This highlights a clear absence of a model capable of processing 3D data…