← Search

Pengzhan Sun

4 accepted papers

2025

Analyzing the Synthetic-to-Real Domain Gap in 3D Hand Pose Estimation

CVPR 2025poster

Recent synthetic 3D human datasets for the face, body, and hands have pushed the limits on photorealism. Face recognition and body pose estimation have achieved state-of-the-art performance using synthetic training data alone, but for the hand, there is still a large synthetic-to-real gap. This pape…

2025

Visual Intention Grounding for Egocentric Assistants

ICCV 2025poster

Visual grounding associates textual descriptions with objects in an image. Conventional methods target third-person image inputs and named object queries. In applications such as AI assistants, the perspective shifts -- inputs are egocentric, and objects may be referred to implicitly through needs a…

2024

ImageCaptioner2: Image Captioner for Image Captioning Bias Amplification Assessment

AAAI 2024technical

Most pre-trained learning systems are known to suffer from bias, which typically emerges from the data, the model, or both. Measuring and quantifying bias and its sources is a challenging task and has been extensively studied in image captioning. Despite the significant effort in this direction, we…

2023

HRS-Bench: Holistic, Reliable and Scalable Benchmark for Text-to-Image Models

ICCV 2023poster

Designing robust text-to-image (T2I) models have been extensively explored in recent years, especially with the emergence of diffusion models, which achieves state-of-the-art results on T2I synthesis tasks. Despite the significant effort and success in this direction, we observed that the existing m…

Cited by 73PDFcodeScholar