← Search

Di Huang*

3 accepted papers

2024

Crowd-SAM:SAM as a smart annotator for object detection in crowded scenes

ECCV 2024poster

"Object detection is an important task that finds its application in a wide range of scenarios. Generally, it requires extensive labels for training, which is quite time-consuming, especially in crowded scenes. Recently, Segment Anything Model (SAM) has emerged as a powerful zero-shot segmenter, off…

2024

Multi-modal Relation Distillation for Unified 3D Representation Learning

ECCV 2024poster

"Recent advancements in multi-modal pre-training for 3D point clouds have demonstrated promising results by aligning heterogeneous features across 3D shapes and their corresponding 2D images and language descriptions. However, current straightforward solutions often overlook intricate structural rel…

Cited by 0SourcePDFScholar
2024

PredBench: Benchmarking Spatio-Temporal Prediction across Diverse Disciplines

ECCV 2024poster

"In this paper, we introduce PredBench, a benchmark tailored for the holistic evaluation of spatio-temporal prediction networks. Despite significant progress in this field, there remains a lack of a standardized framework for a detailed and comparative analysis of various prediction network architec…