← Search

Hongwei Sheng

6 accepted papers

2025

3DRealCar: An In-the-wild RGB-D Car Dataset with 360-degree Views

ICCV 2025poster

3D cars are widely used in self-driving systems, virtual and augmented reality, and gaming applications. However, existing 3D car datasets are either synthetic or low-quality, limiting their practical utility and leaving a significant gap with the high-quality real-world 3D car dataset. In this pape…

Cited by 0SourcePDFScholar
2025

Multimodal Retina Image Analysis Survey: Datasets, Tasks and Methods

IJCAI 2025

Retina images provide a noninvasive view of the central nervous system and microvasculature, making it essential for clinical applications. Changes in the retina often indicate both ophthalmic and systemic diseases, aiding in diagnosis and early intervention. While deep learning algorithms have adva

Cited by 0SourcePDFScholar
2024

Benchmarking Audio Visual Segmentation for Long-Untrimmed Videos

CVPR 2024poster

Existing audio-visual segmentation datasets typically focus on short-trimmed videos with only one pixel-map annotation for a per-second video clip. In contrast for untrimmed videos the sound duration start- and end-sounding time positions and visual deformation of audible objects vary significantly.…

2024

MM-WLAuslan: Multi-View Multi-Modal Word-Level Australian Sign Language Recognition Dataset

NeurIPS 2024poster

Isolated Sign Language Recognition (ISLR) focuses on identifying individual sign language glosses. Considering the diversity of sign languages across geographical regions, developing region-specific ISLR datasets is crucial for supporting communication and research. Auslan, as a sign language specif…

Cited by 0SourcePDFScholar
2023

Auslan-Daily: Australian Sign Language Translation for Daily Communication and News

NeurIPS 2023poster

Sign language translation (SLT) aims to convert a continuous sign language video clip into a spoken language. Considering different geographic regions generally have their own native sign languages, it is valuable to establish corresponding SLT datasets to support related communication and research.…

Cited by 18SourcePDFScholar
2023

RVD: A Handheld Device-Based Fundus Video Dataset for Retinal Vessel Segmentation

NeurIPS 2023poster

Retinal vessel segmentation is generally grounded in image-based datasets collected with bench-top devices. The static images naturally lose the dynamic characteristics of retina fluctuation, resulting in diminished dataset richness, and the usage of bench-top devices further restricts dataset scal…

Cited by 10SourcePDFScholar